Sr. AI infra engineer

Proxima Beta Pte. Limited

Singapore

On-site

SGD 180,000 - 260,000

Full time

41 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Tencent Overseas IT leads global IT infrastructure strategy and delivery, focusing on bridging internal AI infrastructure demand with external resources. You will drive cross-functional cloud/edge projects, turning business needs into deliverable plans, coordinating data centre readiness, and ensuring clear ownership during handover to operations.

The role requires ensuring performance, capacity planning, and governance across compute, network, and storage, with close collaboration to internal

Qualifications

  • Substantial experience in infrastructure architecture or network engineering.
  • Experience delivering or operating networks across data centres or multiple sites.
  • Strong documentation and stakeholder communication skills.

Responsibilities

  • Translate demand into infrastructure requirements with capacity planning and acceptance criteria.
  • Lead network design and delivery reviews across data centres and WAN connectivity.
  • Coordinate network providers, carriers, colocation partners and system integrators on delivery milestones.
  • Drive data centre readiness including rack space, power, cooling, and installation prerequisites.
  • Own integration and acceptance governance across compute, network and storage teams with evidence of performance metrics.
  • Maintain integrated plans, risks, and supplier action trackers with clear handover documentation.

Skills

Infrastructure architecture
Network engineering
Technical delivery
Service management
Data centre networking
TCP/IP & BGP/OSPF
Vendor management

Job description

About the Hiring Team

Tencent Overseas IT has the mission to empower Tencent's rapid global growth with future ready, global IT platforms, applications and services. We are chartered to lead the Overseas IT strategy, architecture, roadmap and execution. Satisfying our internal/external customers and becoming a world class global IT team are our top aspirations.

What the Role Entails

About the Team The AI Compute Centre sits within Tencent's Overseas IT department, acting as the bridge between internal AI infrastructure demand and the external resources that fulfil it. We play key roles in the full lifecycle of AI infrastructure clusters — from requirement gathering and capacity planning, through architectural design and development, to delivery, operations, and DevOps — across regions worldwide.

About the Role

You will lead the cloud/edge cloud infrastructure projects you are responsible for delivering. Your focus is to turn business and technical requirements into deliverable plans, hold suppliers accountable, coordinate data centre readiness, and ensure that new capacity is accepted and handed over to operations with clear ownership. This is a cross-functional role. You will work with specialist engineers on compute, networking, storage, facilities, and security decisions; you are not expected to personally configure every system or run every benchmark. You should, however, be able to challenge assumptions, identify gaps between vendor commitments and delivered outcomes, and drive issues to resolution.

What You Will Do
  • Translate demand into infrastructure requirements. Work with AI business and platform teams to define compute capacity, network bandwidth and latency, connectivity, data access, scalability, availability, and operational needs. Convert these into supplier requirements, evaluation criteria, delivery milestones, and acceptance criteria.
  • Lead network design and delivery reviews. Review data centre fabric, inter-data-centre and WAN connectivity, routing, IP planning, network segmentation, resilience, and expansion plans with internal architects and suppliers. Ensure designs address both AI workload requirements and day-to-day operability.
  • Manage network providers and delivery dependencies. Coordinate network equipment vendors, carriers, colocation partners, and system integrators on circuits, cross-connects, cabling, configuration handoffs, implementation windows, and fault escalation. Track responsibilities and resolve gaps between suppliers.
  • Drive data centre readiness. Coordinate rack space, power, cooling, physical access, cabling, carrier connectivity, and installation prerequisites. Identify site or network dependencies that could delay cluster deployment.
  • Own integration and acceptance governance. Coordinate compute, network, storage, and platform teams through installation, integration, testing, remediation, and sign-off. Require evidence for connectivity, bandwidth, latency, failover, and agreed cluster-level outcomes—not just confirmation that equipment has been installed.
  • Manage delivery and service accountability. Maintain integrated plans, risks, decision logs, and supplier action trackers. Define handover documentation, monitoring expectations, incident escalation, change procedures, support boundaries, and service review cadence.
  • Communicate technical risks clearly. Explain design trade-offs, unresolved defects, capacity constraints, and schedule impacts to both engineering teams and business stakeholders.
Who We Look For
Must-haves
  • Substantial experience in infrastructure architecture, network engineering, technical delivery, or service management, typically 10+ years in relevant roles.
  • Strong data centre or enterprise networking experience, with the ability to review network topology, routing, connectivity, redundancy, and capacity plans—not only manage project schedules.
  • Working knowledge of TCP/IP, L2/L3 networking, VLANs, routing protocols such as BGP and OSPF, and common data centre network design principles.
  • Experience delivering or operating networks across data centres, cloud environments, or multiple sites, including coordination with carriers, colocation providers, and network equipment vendors.
  • Ability to review network test plans and evidence, identify gaps in connectivity or failover, and work with engineers and suppliers to resolve issues.
  • Experience with a data centre migration, site deployment, colocation delivery, or comparable infrastructure transition involving multiple technical teams and external providers.
  • Demonstrated experience managing suppliers through design review, delivery milestones, acceptance, escalation, and production handover.
  • Ability to assess infrastructure proposals across compute, network, storage, availability, and operational support, while engaging specialists for detailed validation.
  • Strong documentation and stakeholder communication skills.
Nice to Have
  • Experience evaluating or sourcing AI server capacity, AI infrastructure services, or large-scale compute platforms.
  • Deeper experience with Spine-Leaf/Clos fabrics, EVPN-VXLAN, ECMP, MLAG, or inter-data-centre connectivity.
  • Familiarity with InfiniBand, RoCE, RDMA, AI server cluster network design, high-performance storage, or cluster-level performance testing such as NCCL.
  • Experience with high-density data centre deployments, including power and cooling constraints.
  • Exposure to supplier SLAs, commercial evaluation, cost optimisation, disaster recovery, or service governance.
  • Relevant networking, cloud, AI infrastructure, project management, or IT service management certifications.
Equal Employment Opportunity at Tencent

As an equal opportunity employer, we firmly believe that diverse voices fuel our innovation and allow us to better serve our users and the community. We foster an environment where every employee of Tencent feels supported and inspired to achieve individual and common goals.

Who we are

Tencent is a world-leading internet and technology company that develops innovative products and services to improve the quality of life for people around the world.

As an equal opportunity employer, we firmly believe that diverse voices fuel our innovation and allow us to better serve our users and the community. We foster an environment where every employee of Tencent feels supported and inspired to achieve individual and common goals.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Compute Intern
AI Compute Intern

Proxima Beta Pte. Limited • Singapore

On-site
SGD 20,000 - 33,000
Software Engineer II
Software Engineer II

Proxima Beta Pte. Limited • Singapore

On-site
SGD 120,000 - 180,000
Senior Cloud & Data Centre Delivery Engineer
Senior Cloud & Data Centre Delivery Engineer

Tencent • Singapore

On-site
SGD 260,000 - 360,000
AI Compute Intern
AI Compute Intern

Tencent • Singapore

On-site
SGD 20,000 - 33,000
Exposure to AI infrastructure
AI IT Engineer Intern
AI IT Engineer Intern

Lightspeed Studios • Singapore

On-site
SGD 12,000 - 18,000
Data Engineer Intern
Data Engineer Intern

Proxima Beta Pte. Limited • Singapore

On-site
SGD 90,000 - 150,000
Senior Data Center Operations Engineer
Senior Data Center Operations Engineer

WeChat International Pte. Ltd. • Singapore

On-site
SGD 90,000 - 170,000
Data Engineer Intern
Data Engineer Intern

Tencent • Singapore

On-site
SGD 110,000 - 180,000
Tencent Cloud – AI & LLM Solution Architecture Leader
Tencent Cloud – AI & LLM Solution Architecture Leader

TENCENT CLOUD INTERNATIONAL PTE. LTD. • Singapore

On-site
SGD 180,000 - 280,000
Senior AI Infrastructure Architect
Senior AI Infrastructure Architect

Proxima Beta Pte. Limited • Singapore

On-site
SGD 180,000 - 260,000