HPC/AI Cluster Engineer (GPU Specialist)

PaleBlueDot AI

Singapore

On-site

SGD 120,000 - 180,000

Full time

10 days ago
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

PaleBlueDot AI in Singapore is seeking an experienced HPC/AI cluster engineer to manage the full lifecycle of data center servers and lead deployment of GPU-accelerated resources. You will optimize performance, ensure stability, and maintain operations for large-scale HPC environments.

The role requires hands-on expertise with MPI/OpenMP, virtualization, benchmarking, and HPC file systems such as Lustre/GPFS, plus strong English communication. Singapore-based on-site position.

Qualifications

  • 5+ years of experience building large-scale HPC clusters.
  • Hands-on GPUs and thousands of GPUs; GPU hardware architecture experience.
  • Proficient in parallel computing technologies MPI and OpenMP.
  • Strong virtualization and containerization knowledge; HPC benchmarking experience (HPL/NCCL).
  • Familiar with high-performance file systems such as Lustre/GPFS.
  • HPC certifications are preferred.
  • Professional working proficiency in English.

Responsibilities

  • Manage the full lifecycle of data center servers, including design, deployment, performance tuning, and validation of HPC/AI clusters.
  • Lead deployment of heterogeneous computing resources (GPUs/XPUs); optimize system performance and stability; monitor and maintain server operations.
  • Prepare technical documentation, provide customer support, and develop automation scripts using Shell, Python, and Ansible.

Skills

GPU architecture
Parallel computing
MPI
OpenMP
HPC clusters
Automation scripting

Education

Bachelor's degree in a related field

Tools

Shell
Python
Ansible

Job description

PaleBlueDot AI in Singapore is seeking an experienced HPC/AI cluster engineer to manage the full lifecycle of data center servers and lead deployment of GPU-accelerated resources. You will optimize performance, ensure stability, and maintain operations for large-scale HPC environments.

The role requires hands-on expertise with MPI/OpenMP, virtualization, benchmarking, and HPC file systems such as Lustre/GPFS, plus strong English communication. Singapore-based on-site position.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

GPU HPC Infra Engineer for AI Clusters
GPU HPC Infra Engineer for AI Clusters

The Supreme HR Advisory Pte. Ltd. • Singapore

On-site
SGD 56,000 - 78,000
AI Infra Engineer: GPU Clusters & HPC Networking
AI Infra Engineer: GPU Clusters & HPC Networking

The Supreme HR Advisory Pte Ltd • Singapore

On-site
SGD 56,000 - 78,000
Server Engineer(AI Cluster/GPU)
Server Engineer(AI Cluster/GPU)

PaleBlueDot AI • Singapore

On-site
SGD 120,000 - 180,000
Senior AI Infra Architect: GPU Clusters & HPC Ops
Senior AI Infra Architect: GPU Clusters & HPC Ops

THE SUPREME HR ADVISORY PTE. LTD. • Singapore

On-site
SGD 56,000 - 78,000
GPU HPC Infrastructure Engineer
GPU HPC Infrastructure Engineer

THE SUPREME HR ADVISORY PTE. LTD. • Singapore

On-site
SGD 56,000 - 78,000
5-day work week
Senior AI Compute & HPC Infrastructure Engineer
Senior AI Compute & HPC Infrastructure Engineer

THE SUPREME HR ADVISORY PTE. LTD. • Singapore

On-site
SGD 56,000 - 78,000
Senior AI/HPC Compute Infrastructure Engineer
Senior AI/HPC Compute Infrastructure Engineer

NVIDIA Gruppe • Singapore

On-site
SGD 120,000 - 190,000
AI Infrastructure Engineer — GPU & HPC Clusters
AI Infrastructure Engineer — GPU & HPC Clusters

RUNSUN SERVICE PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
AI Cloud Network Engineer for HPC & GPU Clusters
AI Cloud Network Engineer for HPC & GPU Clusters

Trulyyy • Singapore

On-site
SGD 120,000 - 180,000
AI HPC Infra Engineer — GPU Clusters & Slurm Expert
AI HPC Infra Engineer — GPU Clusters & Slurm Expert

The Supreme HR Advisory Pte. Ltd. • Singapore

On-site
SGD 56,000 - 78,000