GPU Performance Engineer: Scale ML Inference & Systems

Anthropic

New York (NY)

Hybrid

USD 280,000 - 850,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Anthropic is seeking a GPU Performance Engineer to architect and optimize GPU systems powering Claude and related AI models. You will maximize GPU utilization, implement kernel-level optimizations, and scale pipelines across thousands of GPUs.

Ideal candidates will have deep GPU programming experience, a track record of performance improvements, and enjoy solving complex hardware-to-software challenges in a collaborative environment.

Qualifications

  • Experience with GPU programming and optimization at scale.
  • Ability to navigate complex hardware and ML frameworks.
  • Proven track record delivering measurable performance improvements.
  • Willingness to collaborate and pair program.

Responsibilities

  • Design and optimize GPU kernels for large models.
  • Develop distributed multi-node GPU training and inference pipelines.
  • Profile and optimize GPU utilization in production ML systems.
  • Collaborate with researchers and engineers to push AI infrastructure forward.

Skills

GPU programming
System optimization
Collaborative problem-solving
ML frameworks familiarity

Education

Bachelor’s degree

Tools

CUDA
Triton
CUTLASS
Flash Attention
Nsight

Job description

Anthropic is seeking a GPU Performance Engineer to architect and optimize GPU systems powering Claude and related AI models. You will maximize GPU utilization, implement kernel-level optimizations, and scale pipelines across thousands of GPUs.

Ideal candidates will have deep GPU programming experience, a track record of performance improvements, and enjoy solving complex hardware-to-software challenges in a collaborative environment.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Performance Engineer, GPU
Performance Engineer, GPU

Anthropic • New York (NY)

On-site
USD 280,000 - 850,000
Senior AI Training Performance Engineer (GPU & Scale)
Senior AI Training Performance Engineer (GPU & Scale)

figure.ai • San Jose (CA), Northern (KY)

Hybrid
USD 200,000 - 400,000
Staff Software Engineer, Scalable AI Inference Systems
Staff Software Engineer, Scalable AI Inference Systems

Anthropic • San Francisco (CA)

On-site
USD 320,000 - 485,000
Senior ML Infra Engineer - Scale GPU Clusters, Remote
Senior ML Infra Engineer - Scale GPU Clusters, Remote

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 320,000 - 500,000
Equity
Medical/Dental/Vision coverage
Unlimited PTO
+1
AI Performance Engineer: Scale ML Throughput & Training
AI Performance Engineer: Scale ML Throughput & Training

InvestedintheMission • Sunnyvale (CA)

On-site
USD 180,000 - 260,000
Research Engineer, GPU Performance
Research Engineer, GPU Performance

Harnham • California (MO)

On-site
USD 120,000 - 160,000
Performance Engineer, Inference Engine - Flexible Hours
Performance Engineer, Inference Engine - Flexible Hours

Anthropic • San Francisco (CA), New York (NY)

On-site
USD 350,000 - 850,000
Machine Learning Engineer, GPU Performance
Machine Learning Engineer, GPU Performance

Brahma Consulting Group • San Francisco (CA)

On-site
USD 150,000 - 210,000
GPU Performance Engineer: Accelerate AI Video Pipelines
GPU Performance Engineer: Accelerate AI Video Pipelines

HeyGen • Los Angeles (CA), Palo Alto (CA), San Francisco (CA)

On-site
USD 130,000 - 170,000
Competitive salary
Dynamic work environment
Growth opportunities
+2
Senior ML Systems Engineer – Distributed Training
Senior ML Systems Engineer – Distributed Training

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 280,000
Equity
Health benefits
Remote-friendly US culture
+1