GPU Performance Engineer: Scale ML Inference & Systems

Anthropic

New York (NY)

Hybrid

USD 280,000 - 850,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Anthropic is seeking a GPU Performance Engineer to architect and optimize GPU systems powering Claude and related AI models. You will maximize GPU utilization, implement kernel-level optimizations, and scale pipelines across thousands of GPUs.

Ideal candidates will have deep GPU programming experience, a track record of performance improvements, and enjoy solving complex hardware-to-software challenges in a collaborative environment.

Qualifications

  • Experience with GPU programming and optimization at scale.
  • Ability to navigate complex hardware and ML frameworks.
  • Proven track record delivering measurable performance improvements.
  • Willingness to collaborate and pair program.

Responsibilities

  • Design and optimize GPU kernels for large models.
  • Develop distributed multi-node GPU training and inference pipelines.
  • Profile and optimize GPU utilization in production ML systems.
  • Collaborate with researchers and engineers to push AI infrastructure forward.

Skills

GPU programming
System optimization
Collaborative problem-solving
ML frameworks familiarity

Education

Bachelor’s degree

Tools

CUDA
Triton
CUTLASS
Flash Attention
Nsight

Job description

Anthropic is seeking a GPU Performance Engineer to architect and optimize GPU systems powering Claude and related AI models. You will maximize GPU utilization, implement kernel-level optimizations, and scale pipelines across thousands of GPUs.

Ideal candidates will have deep GPU programming experience, a track record of performance improvements, and enjoy solving complex hardware-to-software challenges in a collaborative environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

GPU Performance Engineer for Scalable AI Inference
GPU Performance Engineer for Scalable AI Inference

SignalAI • New York (NY)

Hybrid
USD 280,000 - 850,000
Competitive compensation
Equity donation matching
Generous vacation and parental leave
+1
Performance Engineer, GPU
Performance Engineer, GPU

Anthropic • New York (NY)

Hybrid
USD 280,000 - 850,000
Senior AI Training Performance Engineer (GPU & Scale)
Senior AI Training Performance Engineer (GPU & Scale)

figure.ai • San Jose (CA), Northern (KY)

Hybrid
USD 200,000 - 400,000
Senior ML Infra Engineer - Scale GPU Clusters, Remote
Senior ML Infra Engineer - Scale GPU Clusters, Remote

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 320,000 - 500,000
Equity
Medical/Dental/Vision coverage
Unlimited PTO
+1
Senior AI Inference Performance Engineer — Scale GPUs
Senior AI Inference Performance Engineer — Scale GPUs

NVIDIA • California (MO)

On-site
USD 124,000 - 196,000
Equity eligibility
ML Performance Engineer: Scale GPU-Driven Training
ML Performance Engineer: Scale GPU-Driven Training

Decisive Point • Sunnyvale (CA)

On-site
USD 180,000 - 240,000
Research Engineer, GPU Performance
Research Engineer, GPU Performance

Harnham • California (MO)

On-site
USD 120,000 - 160,000
GPU Inference Performance Engineer — Equity & Optimization
GPU Inference Performance Engineer — Equity & Optimization

Nvidia Corporation • Santa Clara (CA)

On-site
USD 152,000 - 242,000
Senior GPU Systems Engineer: Scale AI Performance
Senior GPU Systems Engineer: Scale AI Performance

NVIDIA AI • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Benefits
ML Performance Engineer: GPU/CUDA at Scale
ML Performance Engineer: GPU/CUDA at Scale

Selby Jennings • Chicago (IL)

On-site
USD 140,000 - 210,000