AI Systems Engineer: HPC GPU & AIOps

OneAsia Network Limited

Hong Kong

On-site

HKD 420,000 - 640,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Fringe benefits

Job summary

OneAsia Network Limited is seeking an AI System Analyst who sits at the intersection of IT infrastructure, systems engineering, and AIOps to help clients deploy and run AI workloads on GPU clusters.

You will support HPC environments, optimize GPU/INFRA performance, and collaborate with clients to troubleshoot deployment pipelines using Slurm, Kubernetes, Docker, and related tools. A strong Linux/DevOps background plus cloud infra experience is required.

Qualifications

  • Bachelor’s degree or High Diploma in Computer Science, Data Engineering, Information Technology, or relevant field.
  • 2–5 years in Linux systems administration, DevOps, HPC engineering, or cloud infrastructure operations.
  • Deep learning frameworks (PyTorch/TensorFlow), GPU optimization, and Slurm workload orchestration.
  • Python AI ecosystem (essential for debugging AI libraries and automation).
  • Containers & Orchestration: Kubernetes, Docker, Helm charts, and containerized GPU runtime configurations (NVIDIA Container Toolkit).

Responsibilities

  • Support clients in deploying, running, and scaling AI workloads across HPC clusters using Slurm and Kubernetes.
  • Delivers prompt technical support to configure, troubleshoot, and optimize containerized AI microservices (Docker, Kubernetes, vLLM).
  • Assist clients with fine-tuning, training, and serving open-weight and enterprise models (e.g., Qwen, DeepSeek, Llama series).
  • Monitor and maintain GPU/NPU compute clusters (NVIDIA H100/H800) and high-speed interconnects (NVLink, InfiniBand, RoCE).
  • Provide technical guidance to optimize client AI workflows for throughput, memory efficiency, and overall performance.
  • Participate in technical discussions and client meetings to deliver prompt, actionable solutions.

Skills

Linux
DevOps
HPC engineering
Cloud infra
Deep learning
PyTorch
TensorFlow
GPU optimization
Slurm
Kubernetes
Docker
NVIDIA Toolkit
Python

Education

Bachelor's degree
High Diploma

Tools

NVIDIA Container Toolkit

Job description

OneAsia Network Limited is seeking an AI System Analyst who sits at the intersection of IT infrastructure, systems engineering, and AIOps to help clients deploy and run AI workloads on GPU clusters.

You will support HPC environments, optimize GPU/INFRA performance, and collaborate with clients to troubleshoot deployment pipelines using Slurm, Kubernetes, Docker, and related tools. A strong Linux/DevOps background plus cloud infra experience is required.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI System Analyst (Location: Cyberport)
AI System Analyst (Location: Cyberport)

OneAsia Network Limited • Hong Kong

On-site
HKD 420,000 - 640,000
Fringe benefits
Senior AI Infrastructure Architect — GPU Clusters & Security
Senior AI Infrastructure Architect — GPU Clusters & Security

CPJobs International • Hong Kong

On-site
HKD 600,000 - 1,000,000
Principal Core Infrastructure Engineer – GPU
Principal Core Infrastructure Engineer – GPU

Ll Oefentherapie • Hong Kong

On-site
HKD 900,000 - 1,300,000
AI Infrastructure Manager
AI Infrastructure Manager

China Mobile International Limited • Hong Kong

On-site
HKD 800,000 - 1,200,000
AI Transformation & Fullstack Systems Analyst
AI Transformation & Fullstack Systems Analyst

HKT • Hong Kong Island

On-site
HKD 180,000 - 280,000
Senior AI-Powered IT & Network Engineer
Senior AI-Powered IT & Network Engineer

TechJobAsia • Hong Kong Island

On-site
HKD 500,000 - 700,000
Head of AI Infra & HPC Platforms
Head of AI Infra & HPC Platforms

Captiare Limited • Hong Kong Island

On-site
HKD 1,200,000 - 1,800,000
AI Systems Architect/Cloud & AI Infrastructure Lead (over $60K)
AI Systems Architect/Cloud & AI Infrastructure Lead (over $60K)

CPJobs International • Hong Kong

On-site
HKD 600,000 - 1,000,000
Director of AI Infrastructure & HPC Strategy
Director of AI Infrastructure & HPC Strategy

Hong Kong Cyberport Management Co Ltd • Hong Kong

On-site
HKD 600,000 - 1,100,000
Competitive compensation package
HPC Systems Engineer - Hybrid, Cloud & Automation
HPC Systems Engineer - Hybrid, Cloud & Automation

Tower Research Capital • Hong Kong

On-site
HKD 450,000 - 750,000
Generous PTO
Savings plans
Hybrid work
+6