Network Engineer – GPU / AI Infrastructure

VOUCH RECRUITMENT PTE. LTD.

Singapore

On-site

SGD 120,000 - 180,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Our client is a fast-growing AI infrastructure company building next-generation GPU-powered AI platforms. They operate high-performance AI data centers that support large-scale machine learning, AI model training and inference workloads.

The environment is highly technical, focusing on low-latency networking, scalability, automation and operational excellence. The Senior Network Engineer will design, deploy, and manage data center networking for AI workloads, including Spine-Leaf architectures,

Qualifications

  • Bachelor's Degree in CS/CE/IT or a related field.
  • 5+ years designing or supporting enterprise or data center network infrastructure.
  • Strong understanding of modern data center networking.
  • Hands-on experience with BGP, OSPF, VXLAN EVPN, ECMP, MLAG/VPC.
  • Experience with high-speed Ethernet networks (25G/40G/100G/200G/400G).
  • Experience with Cisco Nexus, Arista, Juniper, NVIDIA Spectrum, Mellanox.
  • Familiarity with Linux networking fundamentals.

Responsibilities

  • Design, deploy and support high-performance data center network infrastructure for AI and GPU compute environments.
  • Build, configure, and optimize Spine-Leaf network architectures for large-scale GPU clusters.
  • Deploy and manage high-speed Ethernet and/or InfiniBand fabrics supporting AI/HPC workloads.
  • Configure and maintain Layer 2 and Layer 3 networking technologies, including BGP, OSPF, VXLAN EVPN, ECMP, and MLAG/VPC.
  • Monitor, troubleshoot, and resolve complex network issues across compute, storage, and AI infrastructure.
  • Collaborate closely with Platform, Infrastructure, Linux, DevOps, and AI Engineering teams to deliver scalable AI infrastructure.
  • Develop and maintain network documentation, operational procedures, and technical runbooks.
  • Implement network monitoring, alerting, and automation to improve operational efficiency.
  • Participate in incident response and on-call support for production environments.
  • Evaluate and recommend new networking technologies to improve scalability, reliability, and performance.

Skills

Data center networking
Routing protocols
Network troubleshooting
Cross-functional collaboration
Python scripting
Automation tooling

Education

Bachelor's degree in CS/CE/IT

Tools

Cisco Nexus
Arista
Juniper
NVIDIA Spectrum
Mellanox
InfiniBand

Job description

Our client is a fast-growing AI infrastructure company building next-generation GPU-powered AI platforms. They operate high-performance AI data centers that support large-scale machine learning, AI model training and inference workloads. The environment is highly technical, focusing on low-latency networking, scalability, automation and operational excellence.

Primary Responsibilities
  • Design, deploy and support high-performance data center network infrastructure for AI and GPU compute environments.
  • Build, configure, and optimize Spine-Leaf network architectures for large-scale GPU clusters.
  • Deploy and manage high-speed Ethernet and/or InfiniBand fabrics supporting AI/HPC workloads.
  • Configure and maintain Layer 2 and Layer 3 networking technologies, including BGP, OSPF, VXLAN EVPN, ECMP, and MLAG/VPC.
  • Monitor, troubleshoot, and resolve complex network issues across compute, storage, and AI infrastructure.
  • Optimize network performance, latency, and throughput to support distributed AI training and inference.
  • Perform firmware upgrades, network maintenance, capacity planning, and lifecycle management.
  • Collaborate closely with Platform, Infrastructure, Linux, DevOps, and AI Engineering teams to deliver scalable AI infrastructure.
  • Develop and maintain network documentation, operational procedures, and technical runbooks.
  • Implement network monitoring, alerting, and automation to improve operational efficiency.
  • Participate in incident response and on-call support for production environments.
  • Evaluate and recommend new networking technologies to improve scalability, reliability, and performance.
What We're Looking For
  • Bachelor's Degree in Computer Science, Computer Engineering, Information Technology, or a related discipline.
  • 5+ years of experience designing or supporting enterprise or data center network infrastructure.
  • Strong understanding of modern data center networking principles
  • Hands-on experience with routing and switching protocols such as BGP, OSPF, VXLAN EVPN, ECMP, and MLAG/VPC.
  • Experience managing high-speed Ethernet networks (25G/40G/100G/200G/400G).
  • Exposure to AI, HPC, GPU clusters, or large-scale compute environments would be highly advantageous.
  • Experience working with networking platforms such as Cisco Nexus, Arista, Juniper, NVIDIA Spectrum, or Mellanox.
  • Good understanding of network security concepts including segmentation, ACLs, and firewall policies.
  • Familiarity with Linux networking fundamentals and troubleshooting.
  • Experience with network automation using Python, Ansible, REST APIs, or similar tools is an advantage.
  • Knowledge of technologies such as InfiniBand, RoCEv2, RDMA, GPUDirect, SONiC, or Cumulus Linux is a plus.
  • Strong analytical, troubleshooting, and problem-solving skills.
  • Excellent communication skills with the ability to collaborate effectively across cross-functional engineering teams.
  • Comfortable working in a fast-paced, high-availability production environment supporting mission-critical AI infrastructure.

EA License: 22C1396

EA Personnel: R1551466

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Network Engineer (Data Centre / GPU Infrastructure)
Network Engineer (Data Centre / GPU Infrastructure)

Visa Hunt • Singapore

On-site
SGD 90,000 - 140,000
Senior GPU AI Infrastructure Networking Engineer
Senior GPU AI Infrastructure Networking Engineer

VOUCH RECRUITMENT PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Senior AI Infrastructure & Networking Engineer
Senior AI Infrastructure & Networking Engineer

Genesis Networks Pte Ltd • Singapore

On-site
SGD 180,000 - 240,000
AI Infrastructure Engineer
AI Infrastructure Engineer

The Supreme HR Advisory Pte Ltd • Singapore

On-site
SGD 56,000 - 78,000
Hardware Engineer
Hardware Engineer

RUNSUN SERVICE PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Network Operations Engineer
Network Operations Engineer

RUNSUN SERVICE PTE. LTD. • Singapore

On-site
SGD 90,000 - 130,000
Network Engineer, GPUaaS
Network Engineer, GPUaaS

Singtel Group • Singapore

On-site
SGD 70,000 - 90,000
Full suite of health and wellness benefits
Ongoing training and development programs
Internal mobility opportunities
AB03 - AI Infrastructure Engineer
AB03 - AI Infrastructure Engineer

THE SUPREME HR ADVISORY PTE. LTD. • Singapore

On-site
SGD 56,000 - 78,000
Network Architect
Network Architect

Firmus Technologies • Singapore

On-site
SGD 140,000 - 210,000
Senior Solution Architect, AI Compute Engineer - NVIS
Senior Solution Architect, AI Compute Engineer - NVIS

NVIDIA Gruppe • Singapore

On-site
SGD 120,000 - 190,000