Server Engineer(AI Cluster/GPU)

PaleBlueDot AI

Singapore

On-site

SGD 120,000 - 180,000

Full time

10 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

PaleBlueDot AI in Singapore is seeking an experienced HPC/AI cluster engineer to manage the full lifecycle of data center servers and lead deployment of GPU-accelerated resources. You will optimize performance, ensure stability, and maintain operations for large-scale HPC environments.

The role requires hands-on expertise with MPI/OpenMP, virtualization, benchmarking, and HPC file systems such as Lustre/GPFS, plus strong English communication. Singapore-based on-site position.

Qualifications

  • 5+ years of experience building large-scale HPC clusters.
  • Hands-on GPUs and thousands of GPUs; GPU hardware architecture experience.
  • Proficient in parallel computing technologies MPI and OpenMP.
  • Strong virtualization and containerization knowledge; HPC benchmarking experience (HPL/NCCL).
  • Familiar with high-performance file systems such as Lustre/GPFS.
  • HPC certifications are preferred.
  • Professional working proficiency in English.

Responsibilities

  • Manage the full lifecycle of data center servers, including design, deployment, performance tuning, and validation of HPC/AI clusters.
  • Lead deployment of heterogeneous computing resources (GPUs/XPUs); optimize system performance and stability; monitor and maintain server operations.
  • Prepare technical documentation, provide customer support, and develop automation scripts using Shell, Python, and Ansible.

Skills

GPU architecture
Parallel computing
MPI
OpenMP
HPC clusters
Automation scripting

Education

Bachelor's degree in a related field

Tools

Shell
Python
Ansible

Job description

  • Manage the full lifecycle of data center servers, including design, deployment, performance tuning, and validation, as well as the implementation, operation, and maintenance of HPC/AI clusters.
  • Lead the deployment of heterogeneous computing resources such as GPUs and XPUs, optimize system performance and stability, and monitor and maintain server operations.
  • Prepare technical documentation, provide customer support, and drive the development of automation scripts using Shell, Python, and Ansible.
Key Requirements
  • At least 5 years of relevant experience, with hands‑on expertise in building large‑scale HPC clusters with thousands of GPUs, GPU hardware architecture, and parallel computing technologies such as MPI and OpenMP.
  • Strong proficiency in virtualization and containerization technologies, with experience in HPL and NCCL benchmarking and high‑performance file systems such as Lustre and GPFS.
  • HPC‑related certifications are preferred.
  • Professional working proficiency in English.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Hardware Engineer
Hardware Engineer

RUNSUN SERVICE PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
System Engineer
System Engineer

RUNSUN SERVICE PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
HPC/AI Cluster Engineer (GPU Specialist)
HPC/AI Cluster Engineer (GPU Specialist)

PaleBlueDot AI • Singapore

On-site
SGD 120,000 - 180,000
AI HPC Infra Engineer — GPU Clusters & Slurm Expert
AI HPC Infra Engineer — GPU Clusters & Slurm Expert

The Supreme HR Advisory Pte. Ltd. • Singapore

On-site
SGD 56,000 - 78,000
System Engineer(HPC)
System Engineer(HPC)

RAPSYS TECHNOLOGIES PTE. LTD. • Singapore

On-site
SGD 70,000 - 110,000
Server Engineer
Server Engineer

SUPERPOWER X AI (SINGAPORE) TECHNOLOGY PTE. LTD. • Singapore

On-site
SGD 60,000 - 100,000
AI Systems Infra Engineer - Multi-GPU HPC
AI Systems Infra Engineer - Multi-GPU HPC

THE SUPREME HR ADVISORY PTE. LTD. • Singapore

On-site
SGD 56,000 - 78,000
Senior Systems Engineer
Senior Systems Engineer

fujitsu asia pte ltd • Singapore

On-site
SGD 90,000 - 130,000
Senior Solution Architect, AI Compute Engineer - NVIS
Senior Solution Architect, AI Compute Engineer - NVIS

NVIDIA Gruppe • Singapore

On-site
SGD 120,000 - 190,000
AI Infra Architect: GPU Clusters & HPC
AI Infra Architect: GPU Clusters & HPC

THE SUPREME HR ADVISORY PTE. LTD. • Singapore

On-site
SGD 56,000 - 78,000