Performance Engineer, Inference Engine - Flexible Hours

Anthropic

San Francisco, New York (CA, NY)

On-site

USD 350,000 - 850,000

Full time

7 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Anthropic is seeking a Performance Engineer for the Inference Engine to optimize throughput, latency, and reliability across accelerator and cloud platforms, enabling Claude-scale workloads.

You will model hardware bounds, tune memory, and coordinate host-device interactions, applying deep systems programming to high-performance distributed systems. Collaboration with research, safety, and platform teams is essential.

Qualifications

  • LLM inference concepts including prefill and decode on accelerator compute, memory, interconnect
  • Demonstrated rapid learning in deep, unfamiliar systems and shipped consequential changes quickly
  • Strong systems programming with emphasis on code quality and tests
  • Analytical about performance: observe, profile, hypothesize, test, and measure
  • Low ego: ask questions, take feedback, and assist beyond job scope
  • Enjoy pair programming and collaboration across teams

Skills

LLM inference
Fast learner
Systems programming
Performance analysis
Team collaboration
Pair programming

Tools

Rust

Job description

Anthropic is seeking a Performance Engineer for the Inference Engine to optimize throughput, latency, and reliability across accelerator and cloud platforms, enabling Claude-scale workloads.

You will model hardware bounds, tune memory, and coordinate host-device interactions, applying deep systems programming to high-performance distributed systems. Collaboration with research, safety, and platform teams is essential.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Performance Engineer, Inference Engine - High-Performance AI
Performance Engineer, Inference Engine - High-Performance AI

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Performance Engineer — AI Inference Systems
Performance Engineer — AI Inference Systems

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Visa sponsorship
Flexible hybrid work policy
Performance Engineer, Inference Engine
Performance Engineer, Inference Engine

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Staff Software Engineer, Scaling Inference Systems
Staff Software Engineer, Scaling Inference Systems

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 320,000 - 485,000
Staff Software Engineer, AI Inference Systems
Staff Software Engineer, AI Inference Systems

Anthropic • New York (NY)

Hybrid
USD 320,000 - 485,000
Competitive compensation
Equity donation matching (optional)
Generous vacation
+3
Staff Software Engineer, Scalable AI Inference Systems
Staff Software Engineer, Scalable AI Inference Systems

Anthropic • San Francisco (CA)

On-site
USD 320,000 - 485,000
Senior/Staff Inference Infrastructure Engineer
Senior/Staff Inference Infrastructure Engineer

Menlo Ventures • San Francisco (CA)

On-site
USD 150,000 - 210,000
Staff Software Engineer: Scalable AI Inference Systems
Staff Software Engineer: Scalable AI Inference Systems

Anthropic • Seattle (WA)

Hybrid
USD 320,000 - 485,000
Performance Engineer, Inference Systems San Francisco, CA | New York City, NY | Seattle, WA
Performance Engineer, Inference Systems San Francisco, CA | New York City, NY | Seattle, WA

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Visa sponsorship
Flexible hybrid work policy
GPU Performance Engineer: Scale ML Inference & Systems
GPU Performance Engineer: Scale ML Inference & Systems

Anthropic • New York (NY)

Hybrid
USD 280,000 - 850,000