Performance Engineer, Inference Engine - High-Performance AI

EngineersOfAI

San Francisco, Northern (CA, KY)

Hybrid

USD 180,000 - 240,000

Full time

5 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Anthropic is seeking a Performance Engineer for its Inference Engine to optimize LLM processing across accelerator and cloud platforms. You will work on throughput, latency, and cost, ensuring robust, scalable performance from host to device and across chips.

Experience with Rust/C++ and GPU/accelerator programming is essential. The role emphasizes strong systems thinking, observability, and collaboration in a fast-moving AI safety-focused environment.

Qualifications

  • LLM inference: prefill and decode on accelerator compute.
  • Proven quick learner: ramped fast in deep, unfamiliar systems.
  • Strong systems programming (Rust, C++, or similar).
  • Analytical about performance: observe, hypothesize, test, then measure.
  • Low ego: receptive to feedback and collaboration.
  • Enjoy pair programming and consider societal impact of work.

Responsibilities

  • Keep device utilization high; accelerators should not wait due to overhead.
  • Reuse state to avoid recompute when cheaper than recomputing.
  • Build observability to model impact of potential improvements.
  • Deploy improvements across hosts and devices while maintaining safety.

Skills

LLM inference
Rust
C++
systems programming
performance analysis
pair programming
analytical mindset
memory optimization
GPU/Accelerator programming

Tools

CUDA
RDMA
PCIe

Job description

Anthropic is seeking a Performance Engineer for its Inference Engine to optimize LLM processing across accelerator and cloud platforms. You will work on throughput, latency, and cost, ensuring robust, scalable performance from host to device and across chips.

Experience with Rust/C++ and GPU/accelerator programming is essential. The role emphasizes strong systems thinking, observability, and collaboration in a fast-moving AI safety-focused environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Performance Engineer, Inference Engine - Flexible Hours
Performance Engineer, Inference Engine - Flexible Hours

Anthropic • San Francisco (CA), New York (NY)

On-site
USD 350,000 - 850,000
Performance Engineer — AI Inference Systems
Performance Engineer — AI Inference Systems

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Visa sponsorship
Flexible hybrid work policy
Performance Engineer, Inference Engine
Performance Engineer, Inference Engine

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Performance Engineer: AI Inference & Systems (LLM)
Performance Engineer: AI Inference & Systems (LLM)

RadixArk • Palo Alto (CA)

On-site
USD 120,000 - 160,000
Competitive compensation
Comprehensive health benefits
Flexible work arrangements
INFERENCE OPTIMIZATION ENGINEER
INFERENCE OPTIMIZATION ENGINEER

Up Top • United States

Hybrid
USD 180,000 - 320,000
Inference Systems Performance Engineer
Inference Systems Performance Engineer

Adaption Labs • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 260,000
Annual travel stipend
Lunch stipend
Well-Being benefits
AI Inference Engineer: GPU Performance & Rust Stack
AI Inference Engineer: GPU Performance & Rust Stack

Perplexity • San Francisco (CA)

On-site
USD 140,000 - 210,000
Staff Inference Runtime Architect (Rust/Python)
Staff Inference Runtime Architect (Rust/Python)

Anthropic • New York (NY)

Hybrid
USD 405,000 - 485,000
Generous vacation
Parental leave
Flexible working hours
+1
Staff Engineer, Inference Runtime — High-Performance AI Serving
Staff Engineer, Inference Runtime — High-Performance AI Serving

Anthropic • Seattle (WA)

Hybrid
USD 405,000 - 485,000
Inference Systems Engineer — High-Performance ML Runtime
Inference Systems Engineer — High-Performance ML Runtime

The Consensus • San Jose (CA)

On-site
USD 180,000 - 240,000
Medical/dental/vision benefits
Housing subsidy
Relocation support
+2