Network Engineer (AI DC)

Vouch Recruitment

Singapore

On-site

SGD 120,000 - 180,000

Full time

25 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Vouch Recruitment represents a client that is a fast-growing AI infrastructure company building GPU-powered AI platforms and data centers. You will design and maintain high-performance network infrastructure for AI workloads and collaborate with cross-functional teams to ensure low latency and high reliability.

The role involves Spine-Leaf deployments, Ethernet/InfiniBand fabrics, and automation for scalable AI infrastructure, with on-call and incident support in a highly available production

Qualifications

  • Bachelor's degree in CS/CE/IT or related field.
  • 5+ years in enterprise or data center network design or support.
  • Hands-on with routing protocols such as BGP, OSPF, VXLAN EVPN, ECMP, and MLAG/VPC.
  • Experience with high-speed networks (25G/40G/100G/200G/400G).
  • Knowledge of InfiniBand, RoCEv2, RDMA is a plus.
  • Experience with Cisco Nexus, Arista, Juniper, NVIDIA Spectrum, or Mellanox.

Responsibilities

  • Design, deploy and support high-performance data center network infrastructure for AI and GPU compute environments.
  • Build, configure, and optimize Spine-Leaf network architectures for large-scale GPU clusters.
  • Deploy and manage high-speed Ethernet and/or InfiniBand fabrics supporting AI/HPC workloads.
  • Configure and maintain Layer 2 and Layer 3 networking technologies, including BGP, OSPF, VXLAN EVPN, ECMP, and MLAG/VPC.
  • Monitor, troubleshoot, and resolve complex network issues across compute, storage, and AI infrastructure.
  • Optimize network performance, latency, and throughput to support distributed AI training and inference.
  • Perform firmware upgrades, network maintenance, capacity planning, and lifecycle management.
  • Collaborate with Platform, Infrastructure, Linux, DevOps, and AI Engineering teams.
  • Develop and maintain network documentation, runbooks, and procedures.
  • Implement network monitoring, alerting, and automation to improve operations.
  • Participate in incident response and on-call support for production environments.
  • Evaluate and recommend new networking technologies to improve scalability, reliability, and performance.

Skills

Data center networking
Routing protocols
Network automation
High-speed networks

Education

Bachelor's degree

Tools

Cisco Nexus
Arista
Juniper
NVIDIA Spectrum
Mellanox

Job description

Our client is a fast-growing AI infrastructure company building next-generation GPU-powered AI platforms. They operate high-performance AI data centers that support large-scale machine learning, AI model training and inference workloads. The environment is highly technical, focusing on low-latency networking, scalability, automation and operational excellence.

Primary Responsibilities

  • Design, deploy and support high-performance data center network infrastructure for AI and GPU compute environments.
  • Build, configure, and optimize Spine-Leaf network architectures for large-scale GPU clusters.
  • Deploy and manage high-speed Ethernet and/or InfiniBand fabrics supporting AI/HPC workloads.
  • Configure and maintain Layer 2 and Layer 3 networking technologies, including BGP, OSPF, VXLAN EVPN, ECMP, and MLAG/VPC.
  • Monitor, troubleshoot, and resolve complex network issues across compute, storage, and AI infrastructure.
  • Optimize network performance, latency, and throughput to support distributed AI training and inference.
  • Perform firmware upgrades, network maintenance, capacity planning, and lifecycle management.
  • Collaborate closely with Platform, Infrastructure, Linux, DevOps, and AI Engineering teams to deliver scalable AI infrastructure.
  • Develop and maintain network documentation, operational procedures, and technical runbooks.
  • Implement network monitoring, alerting, and automation to improve operational efficiency.
  • Participate in incident response and on-call support for production environments.
  • Evaluate and recommend new networking technologies to improve scalability, reliability, and performance.

What We're Looking For

  • Bachelor's Degree in Computer Science, Computer Engineering, Information Technology, or a related discipline.
  • 5+ years of experience designing or supporting enterprise or data center network infrastructure.
  • Strong understanding of modern data center networking principles
  • Hands-on experience with routing and switching protocols such as BGP, OSPF, VXLAN EVPN, ECMP, and MLAG/VPC.
  • Experience managing high-speed Ethernet networks (25G/40G/100G/200G/400G).
  • Exposure to AI, HPC, GPU clusters, or large-scale compute environments would be highly advantageous.
  • Experience working with networking platforms such as Cisco Nexus, Arista, Juniper, NVIDIA Spectrum, or Mellanox.
  • Good understanding of network security concepts including segmentation, ACLs, and firewall policies.
  • Familiarity with Linux networking fundamentals and troubleshooting.
  • Experience with network automation using Python, Ansible, REST APIs, or similar tools is an advantage.
  • Knowledge of technologies such as InfiniBand, RoCEv2, RDMA, GPUDirect, SONiC, or Cumulus Linux is a plus.
  • Strong analytical, troubleshooting, and problem-solving skills.
  • Excellent communication skills with the ability to collaborate effectively across cross-functional engineering teams.
  • Comfortable working in a fast-paced, high-availability production environment supporting mission-critical AI infrastructure.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Network Engineer – GPU / AI Infrastructure
Network Engineer – GPU / AI Infrastructure

VOUCH RECRUITMENT PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Network Engineer (Data Centre / GPU Infrastructure)
Network Engineer (Data Centre / GPU Infrastructure)

Visa Hunt • Singapore

On-site
SGD 90,000 - 140,000
Senior AI Infrastructure & Networking Engineer
Senior AI Infrastructure & Networking Engineer

Genesis Networks Pte Ltd • Singapore

On-site
SGD 180,000 - 240,000
Network Operations Engineer
Network Operations Engineer

RUNSUN SERVICE PTE. LTD. • Singapore

On-site
SGD 90,000 - 130,000
Senior DC Network Engineer (Infiniband and AI-fabric )
Senior DC Network Engineer (Infiniband and AI-fabric )

aryan solutions pte. ltd. • Singapore

On-site
SGD 120,000 - 180,000
Senior GPU AI Infrastructure Networking Engineer
Senior GPU AI Infrastructure Networking Engineer

VOUCH RECRUITMENT PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Network Engineer
Network Engineer

SUPERPOWER X AI (SINGAPORE) TECHNOLOGY PTE. LTD. • Singapore

On-site
SGD 70,000 - 120,000
Data Centre Network Engineer – Fabric, Automation & Network AI
Data Centre Network Engineer – Fabric, Automation & Network AI

Jobtailor • Singapore

On-site
SGD 90,000 - 150,000
GPU AI Data Center Networking Engineer
GPU AI Data Center Networking Engineer

Vouch Recruitment • Singapore

On-site
SGD 120,000 - 180,000
AI Infrastructure Engineer
AI Infrastructure Engineer

The Supreme HR Advisory Pte Ltd • Singapore

On-site
SGD 56,000 - 78,000