AI & HPC Systems Lead for GPU Clusters

KLA

Singapore

On-site

SGD 180,000 - 240,000

Full time

12 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

KLA is seeking a System Design Engineering Manager specializing in AI Systems and High‑Performance Computing (HPC) to lead the platform for DL training and inference within the LS-SWIFT division of the Global Products Group. You will guide AI infrastructure design, optimization, and collaboration across hardware and software teams to accelerate innovation.

You will mentor engineers, drive a technical roadmap for DL/AI in HPC, and ensure seamless integration of AI workloads with existing clusters

Qualifications

  • Proven track record leading AI/HPC engineering teams.
  • Familiarity with TensorFlow and PyTorch.
  • Experience with Slurm/OpenMPI.
  • Strong GPU/CUDA programming knowledge.

Responsibilities

  • Lead engineers and system architects to develop the platform for DL training and inference.
  • Define the technical vision, strategy, and roadmap for DL/AI systems within the HPC domain.
  • Collaborate with hardware and software teams to design and optimize AI infrastructure.
  • Ensure seamless integration of AI workloads with existing HPC clusters.
  • Oversee GPU-based cluster management and ensure scalable AI workloads.
  • Benchmark AI models and algorithms on HPC clusters.
  • Communicate technical progress, risks, and opportunities to senior leadership.

Skills

Team management
AI/HPC
TensorFlow/PyTorch
CUDA programming
Docker/Kubernetes

Tools

Slurm
OpenMPI
GPU clusters
Docker
Kubernetes

Job description

KLA is seeking a System Design Engineering Manager specializing in AI Systems and High‑Performance Computing (HPC) to lead the platform for DL training and inference within the LS-SWIFT division of the Global Products Group. You will guide AI infrastructure design, optimization, and collaboration across hardware and software teams to accelerate innovation.

You will mentor engineers, drive a technical roadmap for DL/AI in HPC, and ensure seamless integration of AI workloads with existing clusters

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI & HPC Systems Engineering Manager
AI & HPC Systems Engineering Manager

KLA Corporation • Singapore

On-site
SGD 180,000 - 240,000
AI & HPC Systems Engineering Manager
AI & HPC Systems Engineering Manager

KLA-Belgium • Singapore

On-site
SGD 180,000 - 240,000
Competitive total rewards package
Family friendly
HPC Systems Manager
HPC Systems Manager

KLA • Singapore

On-site
SGD 180,000 - 240,000
HPC Systems Manager
HPC Systems Manager

KLA Corporation • Singapore

On-site
SGD 180,000 - 240,000
HPC Systems Manager
HPC Systems Manager

KLA-Belgium • Singapore

On-site
SGD 180,000 - 240,000
Competitive total rewards package
Family friendly
AI Systems Infra Engineer - Multi-GPU HPC
AI Systems Infra Engineer - Multi-GPU HPC

THE SUPREME HR ADVISORY PTE. LTD. • Singapore

On-site
SGD 56,000 - 78,000
Senior AI Infra Architect: GPU Clusters & HPC Ops
Senior AI Infra Architect: GPU Clusters & HPC Ops

THE SUPREME HR ADVISORY PTE. LTD. • Singapore

On-site
SGD 56,000 - 78,000
AI Infrastructure Engineer - GPU HPC & Kubernetes Expert
AI Infrastructure Engineer - GPU HPC & Kubernetes Expert

THE SUPREME HR ADVISORY PTE. LTD. • Singapore

On-site
SGD 56,000 - 78,000
AI Infra Architect: GPU Clusters & HPC
AI Infra Architect: GPU Clusters & HPC

THE SUPREME HR ADVISORY PTE. LTD. • Singapore

On-site
SGD 56,000 - 78,000
AI Infra Engineer: GPU HPC Clusters & Orchestration
AI Infra Engineer: GPU HPC Clusters & Orchestration

THE SUPREME HR ADVISORY PTE. LTD. • Singapore

On-site
SGD 56,000 - 78,000