AI Infrastructure Engineer - Large-Scale GPU Clusters

SwapeTech

Singapore

On-site

SGD 180,000 - 260,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

SwapeTech in Singapore seeks exceptional AI Engineers to build the next generation of AI infrastructure and MLSys platforms, focusing on large-scale systems rather than model research. You will design, deploy, and operate Kubernetes-based AI infra, RDMA networking, KV Cache memory systems, and CUDA-level performance enhancements.

You will collaborate with AI researchers, infrastructure architects, networking engineers, and platform teams to maximize efficiency, scalability, and reliability

Qualifications

  • Bachelor's degree or higher in Computer Science, Software Engineering, or related fields.
  • Strong software engineering and system design skills.
  • Experience with distributed systems and HPC concepts.
  • Proficiency in C++, Python, or Go.

Responsibilities

  • Design, deploy, and operate large-scale Kubernetes-based AI infrastructure.
  • Develop cluster governance, scheduling, and multi-tenancy.
  • Optimize GPU orchestration and CUDA kernel performance.
  • Improve observability, reliability, and scalability of AI platforms.

Skills

C++
Go
Python
Rust
Distributed Systems
GPU Programming
CUDA

Education

Bachelor's degree in CS/Engineering

Tools

Kubernetes GPU Operator
NVIDIA Network Operator
Prometheus
Grafana
OpenTelemetry
NCCL
UCX

Job description

SwapeTech in Singapore seeks exceptional AI Engineers to build the next generation of AI infrastructure and MLSys platforms, focusing on large-scale systems rather than model research. You will design, deploy, and operate Kubernetes-based AI infra, RDMA networking, KV Cache memory systems, and CUDA-level performance enhancements.

You will collaborate with AI researchers, infrastructure architects, networking engineers, and platform teams to maximize efficiency, scalability, and reliability

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Compute & HPC Infrastructure Engineer
Senior AI Compute & HPC Infrastructure Engineer

THE SUPREME HR ADVISORY PTE. LTD. • Singapore

On-site
SGD 56,000 - 78,000
AI Infra Engineer: HPC GPU Clusters & Kubernetes
AI Infra Engineer: HPC GPU Clusters & Kubernetes

The Supreme HR Advisory Pte. Ltd. • Singapore

On-site
SGD 56,000 - 78,000
AI Systems Infra Engineer - Multi-GPU HPC
AI Systems Infra Engineer - Multi-GPU HPC

THE SUPREME HR ADVISORY PTE. LTD. • Singapore

On-site
SGD 56,000 - 78,000
AI Infrastructure Engineer — GPU & HPC Clusters
AI Infrastructure Engineer — GPU & HPC Clusters

RUNSUN SERVICE PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Staff AI Infrastructure Systems Engineer
Staff AI Infrastructure Systems Engineer

Cloudera • Singapore

On-site
SGD 180,000 - 260,000
Generous PTO Policy
Flexible WFH Policy
Mental & Physical Wellness programs
+2
GPU AI Infrastructure Engineer II
GPU AI Infrastructure Engineer II

Proxima Beta Pte. Limited • Singapore

On-site
SGD 120,000 - 180,000
AI HPC Infra Engineer — GPU Clusters & Slurm Expert
AI HPC Infra Engineer — GPU Clusters & Slurm Expert

The Supreme HR Advisory Pte. Ltd. • Singapore

On-site
SGD 56,000 - 78,000
GPU HPC Infra Engineer for AI Clusters
GPU HPC Infra Engineer for AI Clusters

The Supreme HR Advisory Pte. Ltd. • Singapore

On-site
SGD 56,000 - 78,000
Senior AI Infra Engineer – OpenShift, Kubernetes & GPU
Senior AI Infra Engineer – OpenShift, Kubernetes & GPU

Singtel Group • Singapore

On-site
SGD 180,000 - 240,000
AI Infrastructure Engineer - LCYL
AI Infrastructure Engineer - LCYL

THE SUPREME HR ADVISORY PTE. LTD. • Singapore

On-site
SGD 56,000 - 78,000