Staff Engineer, Inference Runtime — High-Performance AI Serving

Anthropic

Seattle (WA)

Hybrid

USD 405,000 - 485,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Anthropic is looking for a Staff Engineer to lead the Inference Runtime team in Seattle, WA. You will oversee the architecture and roadmap for inference serving systems, ensuring performance and correctness across GPU, TPU, and Trainium platforms.

The ideal candidate will have a strong background in systems engineering and software development. Familiarity with high-performance distributed systems and mentoring experience is crucial. This is a hybrid position with a salary range of $405,000—$485,000.

Qualifications

  • Deep background in systems engineering or ML infrastructure.
  • Real depth in at least one accelerator ecosystem (CUDA/GPU, TPU, or Trainium).
  • Experience with high-performance distributed systems.
  • Track record of using engineering metrics for improvements.
  • Strong communication skills to influence technical direction.

Responsibilities

  • Set technical direction and architecture for the inference serving stack.
  • Own and evolve the runtime’s interfaces and structure.
  • Drive efficient accelerator usage across different platforms.
  • Mentor engineers through design and code reviews.

Skills

Systems engineering
ML infrastructure
Performance profiling
Latency optimization
Software engineering experience
Technical alignment

Education

Bachelor’s degree or equivalent

Tools

CUDA/GPU
TPU
Rust
Python
Kubernetes

Job description

Anthropic is looking for a Staff Engineer to lead the Inference Runtime team in Seattle, WA. You will oversee the architecture and roadmap for inference serving systems, ensuring performance and correctness across GPU, TPU, and Trainium platforms.

The ideal candidate will have a strong background in systems engineering and software development. Familiarity with high-performance distributed systems and mentoring experience is crucial. This is a hybrid position with a salary range of $405,000—$485,000.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff Inference Runtime Architect (Rust/Python)
Staff Inference Runtime Architect (Rust/Python)

Anthropic • New York (NY)

Hybrid
USD 405,000 - 485,000
Generous vacation
Parental leave
Flexible working hours
+1
Staff Inference Engineer — Production-Scale AI Platform
Staff Inference Engineer — Production-Scale AI Platform

Designworks Talent LLC • Bellevue (KY)

Hybrid
USD 180,000 - 240,000
Health insurance
401(k) plan with company match
Paid holidays
Senior AI Inference Runtime Architect
Senior AI Inference Runtime Architect

Arm • Seattle (WA)

Hybrid
USD 263,000 - 355,000
Hybrid working
Accommodations during recruitment
Senior Staff Engineer, AI Inference Systems
Senior Staff Engineer, AI Inference Systems

Anthropic • New York (NY)

Hybrid
USD 320,000 - 485,000
Staff Engineer, Inference Tooling & DevX
Staff Engineer, Inference Tooling & DevX

anthropic • San Francisco (CA), Seattle (WA), New York (NY)

Hybrid
USD 405,000 - 485,000
Performance Engineer — AI Inference Systems
Performance Engineer — AI Inference Systems

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Visa sponsorship
Flexible hybrid work policy
Performance Engineer, Inference Engine - High-Performance AI
Performance Engineer, Inference Engine - High-Performance AI

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Senior AI Inference Deployment Engineer
Senior AI Inference Deployment Engineer

Anthropic • Seattle (WA)

Hybrid
USD 320,000 - 485,000
Competitive compensation
Flexible working hours
Generous vacation and parental leave
Software Engineer, Inference - Performance Optimization
Software Engineer, Inference - Performance Optimization

OpenAI, Inc. • San Francisco (CA)

On-site
USD 295,000 - 555,000
Equity
Engineering Manager: Inference Infrastructure Leader
Engineering Manager: Inference Infrastructure Leader

EngineersOfAI • New York (NY), Northern (KY)

Hybrid
USD 230,000 - 360,000