Technical Support Engineer (GPU/HPC)

KERRY CONSULTING PTE. LTD.

Singapore

On-site

SGD 90,000 - 130,000

Full time

3 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Kerry Consulting PTE. LTD. is partnering with a rapidly expanding AI infrastructure and cloud computing company to hire a Senior Technical Support Engineer for GPU & Cloud Infrastructure in Singapore.

This senior role focuses on resolving complex customer and production issues across GPU compute environments while collaborating with Engineering and SRE teams to strengthen platform reliability. You will handle end-to-end escalations, contribute to post-incident reviews, and mentor junior

Qualifications

  • 6+ years of experience in cloud infrastructure technical support, escalation engineering, production operations, or an SRE-adjacent role.
  • Hands‑on expertise across Linux, networking, and cloud infrastructure.
  • Experience with GPU infrastructure, CUDA, HPC, or large‑scale compute environments is highly relevant.

Responsibilities

  • Own complex escalations and platform incidents end to end, troubleshooting across GPU compute, Linux, networking/SDN, storage, CUDA, and drivers.
  • Lead root-cause analysis and drive permanent solutions with engineering teams.
  • Improve monitoring, processes, and runbooks; mentor L1/L2 engineers and contribute to technical documentation.
  • Participate in on-call escalation rotations and incident reviews.

Skills

Cloud infrastructure
Linux
Networking
SRE practices
Root-cause analysis
Communication

Tools

CUDA
GPU compute

Job description

Overview

I'm partnering a rapidly expanding AI infrastructure and cloud computing company to hire a Senior Technical Support Engineer, GPU & Cloud Infrastructurein Singapore. This is a senior technical role focused on resolving complex customer and production issues across GPU compute environments while working closely with Engineering and SRE teams to strengthen platform reliability.

Responsibilities

You will take end-to-end ownership of complex escalations and platform incidents, troubleshooting across GPU compute, Linux, networking/SDN, storage, CUDA and drivers, control plane, and billing systems. Working closely with SRE, Compute, and Engineering teams, you will lead root-cause analysis, identify permanent solutions to recurring issues, and translate operational learnings into stronger monitoring, processes, and runbooks. You will also play an important role during major incidents and post-incident reviews, while mentoring L1/L2 engineers and improving the quality of technical documentation and escalation practices. The position participates in an on-call escalation rotation.

Requirements

You should have at least 6 years of experience in cloud infrastructure technical support, escalation engineering, production operations, or an SRE-adjacent role, with strong hands‑on expertise across Linux, networking, and cloud infrastructure. Experience supporting GPU infrastructure, CUDA, HPC, or large‑scale compute environments will be particularly relevant. You will bring a strong track record of diagnosing difficult production issues, performing structured root‑cause analysis, and collaborating effectively with engineering teams to drive issues through to resolution.

Strong written and verbal communication skills, with the ability to communicate clearly across technical teams, are important,along with comfort handling production incidents and participating in an on-call environment.

License No: 16S8060 / Registration No: R26161178

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Technical Support Engineer (GPU/HPC)
Technical Support Engineer (GPU/HPC)

Kerry Consulting • Singapore

On-site
SGD 120,000 - 180,000
Senior GPU & Cloud Infrastructure Support Engineer
Senior GPU & Cloud Infrastructure Support Engineer

KERRY CONSULTING PTE. LTD. • Singapore

On-site
SGD 90,000 - 130,000
Senior GPU & Cloud Infrastructure Support Engineer
Senior GPU & Cloud Infrastructure Support Engineer

Kerry Consulting • Singapore

On-site
SGD 120,000 - 180,000
DevOps Engineer, GPUaaS
DevOps Engineer, GPUaaS

Singapore Telecommunications Limited • Singapore

On-site
SGD 120,000 - 180,000
Senior AI Infrastructure Reliability Engineer
Senior AI Infrastructure Reliability Engineer

nscale operations apac pte. ltd. • Singapore

On-site
SGD 120,000 - 180,000
Senior AI Infrastructure Support Engineer
Senior AI Infrastructure Support Engineer

nscale operations apac pte. ltd. • Singapore

On-site
SGD 120,000 - 180,000
6723 - GPU Infrastructure Engineer | Up to $7K | Kaki Bukit | NVIDIA, CUDA & HPC
6723 - GPU Infrastructure Engineer | Up to $7K | Kaki Bukit | NVIDIA, CUDA & HPC

The Supreme HR Advisory Pte. Ltd. • Singapore

On-site
SGD 56,000 - 78,000
GPUaaS Network Engineer for AI & HPC
GPUaaS Network Engineer for AI & HPC

Singtel • Singapore

On-site
Confidential
Senior IT Infrastructure & Systems Support Specialist (Singapore)
Senior IT Infrastructure & Systems Support Specialist (Singapore)

Horizon Quantum • Singapore

On-site
SGD 90,000 - 150,000
GPU HPC Infra Engineer for AI Clusters
GPU HPC Infra Engineer for AI Clusters

The Supreme HR Advisory Pte. Ltd. • Singapore

On-site
SGD 56,000 - 78,000