TPU Kernel Engineer: Optimize ML Systems at Scale

Anthropic

York and North Yorkshire

On-site

GBP 110,000 - 150,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Health insurance
Dental & vision
Parental leave 22 weeks
Flexible PTO
Mental health support
Equity package
Relocation assistance
Home office stipend
Meals in office
Education stipend

Job summary

Anthropic in the United Kingdom seeks a TPU Kernel Engineer to identify and address performance issues across ML systems, including research, training and inference. You will design and optimize kernels for the TPU and provide feedback to researchers on how model changes impact performance.

You will implement low-latency, high-throughput sampling for large language models, adapt models for low-precision inference, and build quantitative models of system performance.

Qualifications

  • Bachelor’s degree or equivalent experience required.
  • Strong track record solving large-scale systems optimization.
  • Experience optimizing ML systems on TPUs, GPUs or accelerators.
  • Enjoy pair programming and collaboration.
  • Understand accelerators and computer architecture.
  • Knowledge of ML framework internals.
  • Interest in ML research and societal impact.

Responsibilities

  • Identify and address performance issues across ML systems (research, training, inference).
  • Design and optimize kernels for the TPU.
  • Provide feedback to researchers on model changes and performance impact.
  • Implement low-latency, high-throughput sampling for large language models.
  • Adapt existing models for low-precision inference.
  • Build quantitative models of system performance.
  • Design and implement custom collective communication algorithms.
  • Debug kernel performance at the assembly level.

Skills

Large-scale systems
Low-level optimization
TPU optimization
GPUs & accelerators
Pair programming
Computer architecture
ML framework internals
Transformers / LMs
Societal impact awareness

Education

Bachelor's degree in related field

Job description

Anthropic in the United Kingdom seeks a TPU Kernel Engineer to identify and address performance issues across ML systems, including research, training and inference. You will design and optimize kernels for the TPU and provide feedback to researchers on how model changes impact performance.

You will implement low-latency, high-throughput sampling for large language models, adapt models for low-precision inference, and build quantitative models of system performance.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

TPU Kernel Engineer
TPU Kernel Engineer

Anthropic • York and North Yorkshire

On-site
GBP 110,000 - 150,000
Health insurance
Dental & vision
Parental leave 22 weeks
+7
High-Throughput ML Systems Performance Engineer
High-Throughput ML Systems Performance Engineer

Anthropic • York and North Yorkshire

On-site
GBP 110,000 - 150,000
Health insurance
Fertility benefits
Parental leave 22 weeks
+12
ML Systems Performance Engineer
ML Systems Performance Engineer

Quant Blueprint LLC • Greater London

On-site
GBP 50,000 - 70,000
Remote Performance Engineer: ML Training & Kernels
Remote Performance Engineer: ML Training & Kernels

Cohere • Greater London

On-site
GBP 75,000 - 95,000
Co-working benefit
Daily lunch program
Regular community and social events
Performance Engineer (GPU)
Performance Engineer (GPU)

Anthropic • York and North Yorkshire

On-site
GBP 90,000 - 140,000
Comprehensive health insurance
Fertility benefits
22 weeks parental leave
+1
AI Hardware Acceleration Engineer for ML Performance
AI Hardware Acceleration Engineer for ML Performance

XTX Markets • Greater London

On-site
GBP 60,000 - 100,000
Onsite gym
Extensive medical benefits
Daily breakfast and lunch
+2
Real-Time ML Systems Performance Engineer
Real-Time ML Systems Performance Engineer

Janestreet • Greater London

On-site
GBP 70,000 - 90,000
AI Inference Optimizations Engineer (MLIR/LLVM)
AI Inference Optimizations Engineer (MLIR/LLVM)

microTECH Global Limited • City of Edinburgh

Hybrid
GBP 90,000 - 120,000
Engineering Manager, GPU & ML Systems Scaling
Engineering Manager, GPU & ML Systems Scaling

Anthropic • York and North Yorkshire

On-site
GBP 110,000 - 150,000
Health insurance
Dental insurance
Vision insurance
+15
ML Performance Engineer – Scale GPU/CPU Workloads
ML Performance Engineer – Scale GPU/CPU Workloads

Barlowe LLP • Greater London

On-site
GBP 90,000 - 150,000
Lunch provided
35 days’ annual leave
9% company pension contributions
+4