GPU Performance Engineer: Scale Inference & Training

Anthropic

York and North Yorkshire

On-site

GBP 90,000 - 140,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Comprehensive health insurance
Fertility benefits
22 weeks parental leave
Flexible paid time off

Job summary

Anthropic seeks a GPU Performance Engineer to architect and implement foundational systems powering Claude, driving GPU utilization and optimization at unprecedented scale. You will work at the hardware-software boundary, from custom kernel development to distributed multi-node clusters.

Responsibilities include end-to-end optimization of training and inference pipelines, co-design of attention mechanisms for future architectures, and collaboration with hardware vendors to shape accelerator

Qualifications

  • Track record delivering GPU performance improvements in production ML systems.
  • Experience optimizing end-to-end training and inference pipelines at scale.
  • Strong understanding of modern ML frameworks and low-level GPU hardware.

Responsibilities

  • Architect and implement foundational GPU performance systems for Claude.
  • Maximize GPU utilization and inference efficiency at scale.
  • Develop kernel-level optimizations and distributed architectures across thousands of GPUs.
  • Co-design attention mechanisms for next-gen hardware architectures.
  • Collaborate with hardware vendors to influence accelerator capabilities.

Skills

GPU performance optimization
Collaborative problem-solving
Hardware-software co-design

Education

Bachelor's degree or equivalent experience

Tools

CUDA
Triton
CUTLASS
Flash Attention
Tensor cores
PyTorch/JAX internals
torch.compile
XLA
NCCL
NVLink

Job description

Anthropic seeks a GPU Performance Engineer to architect and implement foundational systems powering Claude, driving GPU utilization and optimization at unprecedented scale. You will work at the hardware-software boundary, from custom kernel development to distributed multi-node clusters.

Responsibilities include end-to-end optimization of training and inference pipelines, co-design of attention mechanisms for future architectures, and collaboration with hardware vendors to shape accelerator

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Performance Engineer (GPU)
Performance Engineer (GPU)

Anthropic • York and North Yorkshire

On-site
GBP 90,000 - 140,000
Comprehensive health insurance
Fertility benefits
22 weeks parental leave
+1
Engineering Manager, GPU & ML Systems Scaling
Engineering Manager, GPU & ML Systems Scaling

Anthropic • York and North Yorkshire

On-site
GBP 110,000 - 150,000
Health insurance
Dental insurance
Vision insurance
+15
Senior AI Infrastructure Performance Engineer - GPU/Inference
Senior AI Infrastructure Performance Engineer - GPU/Inference

Pure Resourcing Solutions • Cambridge

Hybrid
GBP 90,000 - 120,000
ML Performance Engineer – Scale GPU/CPU Workloads
ML Performance Engineer – Scale GPU/CPU Workloads

Barlowe LLP • Greater London

On-site
GBP 90,000 - 150,000
Lunch provided
35 days’ annual leave
9% company pension contributions
+4
Engineering Manager (GPU, ML Accelerator)
Engineering Manager (GPU, ML Accelerator)

Anthropic • York and North Yorkshire

On-site
GBP 110,000 - 150,000
Health insurance
Dental insurance
Vision insurance
+15
Lead AI Training Infrastructure Engineer
Lead AI Training Infrastructure Engineer

Genesis • Greater London

Hybrid
GBP 90,000 - 130,000
GPU Infra Engineer — Scale, Automation & AI Compute
GPU Infra Engineer — Scale, Automation & AI Compute

OpenAI • Greater London

On-site
GBP 120,000 - 190,000
GPU Infrastructure Lead: Scale, Certification, and Automation
GPU Infrastructure Lead: Scale, Certification, and Automation

Hamilton Barnes Associates Limited • Greater London

On-site
GBP 140,000 - 170,000
Full Benefits
Lead GPU Infrastructure Architect for Scalable AI Clusters
Lead GPU Infrastructure Architect for Scalable AI Clusters

Hamilton Barnes Associates Limited • Greater London

On-site
GBP 110,000 - 150,000
AI Compiler Engineer for Next-Gen GPU Inference
AI Compiler Engineer for Next-Gen GPU Inference

NVIDIA • Otley

On-site
GBP 90,000 - 150,000
Generous benefits package