Real-Time AI Systems Engineer

Harnham

United States

On-site

USD 120,000 - 150,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Harnham is looking for a technical expert to accelerate AI systems performance for next-generation models. The role involves optimising GPU training throughput, implementing advanced techniques, and designing scalable systems. Candidates should have over 4 years of relevant experience in performance optimisation, distributed systems, and strong GPU programming skills. This is a unique opportunity to work with cutting-edge AI technology and contribute significantly to real-time systems.

Qualifications

  • 4+ years of experience in systems engineering, ML infrastructure, or performance optimisation.
  • Strong experience with GPU programming.
  • Proven experience building scalable, fault-tolerant training systems.

Responsibilities

  • Optimize training throughput across large GPU clusters.
  • Implement mixed precision and memory-efficient techniques.
  • Design and scale distributed training systems.
  • Profile and optimise inference pipelines for real-time multimodal generation.

Skills

GPU programming (CUDA, Triton)
Performance optimisation
Distributed systems
ML framework internals (PyTorch, JAX)
Mixed or low‑precision techniques (FP8, INT8, BF16)

Job description

Harnham is looking for a technical expert to accelerate AI systems performance for next-generation models. The role involves optimising GPU training throughput, implementing advanced techniques, and designing scalable systems. Candidates should have over 4 years of relevant experience in performance optimisation, distributed systems, and strong GPU programming skills. This is a unique opportunity to work with cutting-edge AI technology and contribute significantly to real-time systems.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Real-Time GPU Performance Engineer for Scalable AI
Real-Time GPU Performance Engineer for Scalable AI

Harnham • California (MO)

On-site
USD 120,000 - 160,000
AI/ML Engineer: Next‑Gen Platforms & GPU Workloads
AI/ML Engineer: Next‑Gen Platforms & GPU Workloads

VeeAR Projects Inc. • Sunnyvale (CA)

On-site
USD 140,000 - 210,000
Senior GPU Performance Engineer for AI Training
Senior GPU Performance Engineer for AI Training

CareerArc • San Jose (CA)

Hybrid
USD 150,000 - 200,000
Competitive salary
Comprehensive benefits
Senior AI Infrastructure Engineer – GPU Systems
Senior AI Infrastructure Engineer – GPU Systems

Clockwork Systems, Inc. • Palo Alto (CA)

On-site
USD 130,000 - 180,000
Competitive compensation
Great benefits package
Catered lunch
Senior AI Performance Engineer — Remote
Senior AI Performance Engineer — Remote

OpenAI • Los Angeles (CA)

On-site
USD 130,000 - 180,000
Research Engineer, GPU Performance
Research Engineer, GPU Performance

Harnham • California (MO)

On-site
USD 120,000 - 160,000
Senior AI Training Performance Engineer (GPU & Scale)
Senior AI Training Performance Engineer (GPU & Scale)

figure.ai • San Jose (CA), Northern (KY)

Hybrid
USD 200,000 - 400,000
Staff Engineer: GPU Kernels & AI Performance
Staff Engineer: GPU Kernels & AI Performance

Gimlet Labs • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior AI Systems Performance Engineer
Senior AI Systems Performance Engineer

NVIDIA • Town of Texas (WI)

On-site
USD 184,000 - 287,500
Equity
Benefits
Remote AI Performance Engineer - Max Throughput
Remote AI Performance Engineer - Max Throughput

Bazeta • Northern (KY)

Hybrid
USD 90,000 - 110,000