AI Infrastructure Engineer - Large-Scale GPU Clusters

SwapeTech

Singapore

On-site

SGD 180,000 - 260,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

SwapeTech in Singapore seeks exceptional AI Engineers to build the next generation of AI infrastructure and MLSys platforms, focusing on large-scale systems rather than model research. You will design, deploy, and operate Kubernetes-based AI infra, RDMA networking, KV Cache memory systems, and CUDA-level performance enhancements.

You will collaborate with AI researchers, infrastructure architects, networking engineers, and platform teams to maximize efficiency, scalability, and reliability

Qualifications

  • Bachelor's degree or higher in Computer Science, Software Engineering, or related fields.
  • Strong software engineering and system design skills.
  • Experience with distributed systems and HPC concepts.
  • Proficiency in C++, Python, or Go.

Responsibilities

  • Design, deploy, and operate large-scale Kubernetes-based AI infrastructure.
  • Develop cluster governance, scheduling, and multi-tenancy.
  • Optimize GPU orchestration and CUDA kernel performance.
  • Improve observability, reliability, and scalability of AI platforms.

Skills

C++
Go
Python
Rust
Distributed Systems
GPU Programming
CUDA

Education

Bachelor's degree in CS/Engineering

Tools

Kubernetes GPU Operator
NVIDIA Network Operator
Prometheus
Grafana
OpenTelemetry
NCCL
UCX

Job description

SwapeTech in Singapore seeks exceptional AI Engineers to build the next generation of AI infrastructure and MLSys platforms, focusing on large-scale systems rather than model research. You will design, deploy, and operate Kubernetes-based AI infra, RDMA networking, KV Cache memory systems, and CUDA-level performance enhancements.

You will collaborate with AI researchers, infrastructure architects, networking engineers, and platform teams to maximize efficiency, scalability, and reliability

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Infrastructure Engineer — GPU & HPC Clusters
AI Infrastructure Engineer — GPU & HPC Clusters

RUNSUN SERVICE PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
AI Systems Engineer: GPU & HPC Clusters
AI Systems Engineer: GPU & HPC Clusters

runsun cloud pte ltd • Singapore

On-site
SGD 120,000 - 180,000
AI GPU Cluster Engineer
AI GPU Cluster Engineer

runsun cloud pte ltd • Singapore

On-site
SGD 100,000 - 160,000
Infra Engineer: GPU Clusters & Kubernetes (Singapore)
Infra Engineer: GPU Clusters & Kubernetes (Singapore)

Exa • Singapore

On-site
SGD 90,000 - 150,000
Senior AI/HPC Systems Engineer
Senior AI/HPC Systems Engineer

NVIDIA • Singapore

On-site
SGD 120,000 - 180,000
GPU AI/HPC DevOps Engineer: Build & Auto-Scale Clusters
GPU AI/HPC DevOps Engineer: Build & Auto-Scale Clusters

Singtel • Singapore

On-site
Confidential
Senior AI Platform Engineer — Kubernetes & GPU Infra
Senior AI Platform Engineer — Kubernetes & GPU Infra

Bitdeer Group • Singapore

On-site
SGD 120,000 - 180,000
Senior AI Platform Engineer — Enterprise AI & GPU Infra
Senior AI Platform Engineer — Enterprise AI & GPU Infra

DSTA • Singapore

On-site
SGD 65,000 - 85,000
AI Engineer (ML Systems & Infrastructure)
AI Engineer (ML Systems & Infrastructure)

SwapeTech • Singapore

On-site
SGD 180,000 - 260,000
GPU Cloud & AI Infrastructure Lead
GPU Cloud & AI Infrastructure Lead

zy future international pte. ltd. • Singapore

On-site
SGD 167,000 - 223,000