Performance Engineer — AI Inference Systems

Anthropic

San Francisco (CA)

Hybrid

USD 350,000 - 850,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Visa sponsorship
Flexible hybrid work policy

Job summary

Anthropic is looking for a Performance Engineer to work on its Inference System Dynamics team. You will focus on optimizing throughput, latency, and correctness across systems serving millions of users. You will also improve the evaluation pipeline for model outputs and build observability tools for performance metrics.

The ideal candidate has hands-on experience in performance engineering, strong data analysis skills, and proficiency in Python. This role offers a hybrid work environment and is integral to maintaining the high standards of AI systems.

Qualifications

  • 1+ years of hands-on performance engineering experience in complex production systems.
  • Ability to read and contribute to large production Python codebases.
  • Solid experience with data analysis tools like SQL or pandas.

Responsibilities

  • Run performance investigations across throughput, latency, and reliability.
  • Improve correctness evaluation pipeline for model output quality.
  • Build observability and dashboards for performance metrics.

Skills

Performance engineering experience
Proficiency in Python
Data analysis skills
Communication of quantitative results
Interest in correctness

Education

Bachelor’s degree or equivalent

Job description

Anthropic is looking for a Performance Engineer to work on its Inference System Dynamics team. You will focus on optimizing throughput, latency, and correctness across systems serving millions of users. You will also improve the evaluation pipeline for model outputs and build observability tools for performance metrics.

The ideal candidate has hands-on experience in performance engineering, strong data analysis skills, and proficiency in Python. This role offers a hybrid work environment and is integral to maintaining the high standards of AI systems.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Performance Engineer, Inference Engine - High-Performance AI
Performance Engineer, Inference Engine - High-Performance AI

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Performance Engineer, Inference Systems San Francisco, CA | New York City, NY | Seattle, WA
Performance Engineer, Inference Systems San Francisco, CA | New York City, NY | Seattle, WA

Anthropic • San Francisco (CA)

On-site
USD 350,000 - 850,000
Visa sponsorship
Flexible hybrid work policy
Performance Engineer, Inference Engine - Flexible Hours
Performance Engineer, Inference Engine - Flexible Hours

Anthropic • San Francisco (CA), New York (NY)

On-site
USD 350,000 - 850,000
Staff Engineer, Inference Runtime — High-Performance AI Serving
Staff Engineer, Inference Runtime — High-Performance AI Serving

Anthropic • Seattle (WA)

Hybrid
USD 405,000 - 485,000
Inference Performance Engineer — Applied AI
Inference Performance Engineer — Applied AI

CoreWeave • Seattle (WA)

On-site
USD 188,000 - 275,000
Medical, dental, and vision insurance
Equity awards
Discretionary bonus
+6
Software Engineer, Inference - Performance Optimization
Software Engineer, Inference - Performance Optimization

OpenAI, Inc. • San Francisco (CA)

On-site
USD 295,000 - 555,000
Equity
Staff Software Engineer: Scalable AI Inference Systems
Staff Software Engineer: Scalable AI Inference Systems

Anthropic • Seattle (WA)

Hybrid
USD 320,000 - 485,000
Performance Engineer, Inference Engine
Performance Engineer, Inference Engine

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Staff Software Engineer, GenAI Inference Performance
Staff Software Engineer, GenAI Inference Performance

Google LLC • Mountain View (CA)

On-site
USD 207,000 - 300,000
AI Performance Engineer – HPC, ARM & Distributed Inference
AI Performance Engineer – HPC, ARM & Distributed Inference

EngineersOfAI • Austin (TX)

On-site
USD 90,000 - 120,000