Real-Time GPU Performance Engineer for Scalable AI

Harnham

California (MO)

On-site

USD 120,000 - 160,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Harnham is seeking a systems engineer to optimize large-scale AI systems for real-time performance. You will work on cutting-edge multimodal models, improving training efficiency and system architecture.

This role requires extensive experience in GPU programming and distributed systems, where you will directly influence AI advancements. Join Harnham to contribute to the future of interactive AI systems.

Qualifications

  • 4+ years of experience in systems engineering, ML infrastructure, or performance optimization.
  • Strong experience with GPU programming (CUDA, Triton, or similar).
  • Experience with distributed systems and large-scale training (NCCL, model parallelism).

Responsibilities

  • Optimize training throughput across large GPU clusters.
  • Implement techniques such as mixed precision, memory-efficient attention, and checkpointing.
  • Profile and optimize inference pipelines for real-time multimodal generation.

Skills

GPU programming
Performance optimization
Distributed systems
Machine Learning frameworks
Low-precision techniques

Job description

Harnham is seeking a systems engineer to optimize large-scale AI systems for real-time performance. You will work on cutting-edge multimodal models, improving training efficiency and system architecture.

This role requires extensive experience in GPU programming and distributed systems, where you will directly influence AI advancements. Join Harnham to contribute to the future of interactive AI systems.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Real-Time AI Systems Engineer
Real-Time AI Systems Engineer

Harnham • United States

On-site
USD 120,000 - 150,000
Remote AI Systems Engineer — Scalable GPU Infra
Remote AI Systems Engineer — Scalable GPU Infra

Bright Vision Technologies • United States

Remote
USD 90,000 - 100,000
Research Engineer, GPU Performance
Research Engineer, GPU Performance

Harnham • California (MO)

On-site
USD 120,000 - 160,000
Senior AI Training Performance Engineer (GPU & Scale)
Senior AI Training Performance Engineer (GPU & Scale)

figure.ai • San Jose (CA), Northern (KY)

Hybrid
USD 200,000 - 400,000
GPU Infrastructure Engineer — Scalable AI Training
GPU Infrastructure Engineer — Scalable AI Training

Thinking Machines Lab Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 350,000 - 475,000
Health benefits
Unlimited PTO
Parental leave
+1
Senior AI Infrastructure Engineer — Scale GPU Clusters Remote
Senior AI Infrastructure Engineer — Scale GPU Clusters Remote

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 280,000 - 420,000
Equity options
Health, vision, dental benefits
Unlimited PTO
+2
Remote GPU Systems Engineer for AI & HPC Optimization
Remote GPU Systems Engineer for AI & HPC Optimization

Bright Vision Technologies • Plymouth (MN)

Remote
USD 100,000 - 150,000
GPU Performance Engineer: Scale ML Inference & Systems
GPU Performance Engineer: Scale ML Inference & Systems

Anthropic • New York (NY)

Hybrid
USD 280,000 - 850,000
AI & HPC GPU Compute Performance Engineer
AI & HPC GPU Compute Performance Engineer

engineeringjobs.net, Inc. • San Jose (CA)

On-site
USD 150,000 - 190,000
Medical Insurance
Dental Insurance
Vision Insurance
+16
Research Engineer
Research Engineer

Harnham • United States

On-site
USD 120,000 - 150,000