Engineering Manager - AI Infrastructure & Platforms

Anthropic

Seattle (WA)

Hybrid

USD 405,000 - 625,000

Full time

3 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Competitive compensation
Equity matching
Parental leave
Flexible hours
Office space

Job summary

Anthropic is seeking an Engineering Manager to lead the inference platform team responsible for the Claude request path and fleet coordination. You will guide a group of ML platform, infrastructure, and distributed-systems engineers to ensure throughput, reliability, and efficient scaling as models, hardware, and clouds evolve.

You will own capacity planning, latency optimization, and cross-team collaboration with product, inference engine, and cloud teams, while driving architectural decisions

Qualifications

  • Engineering management experience leading teams on critical-path production infrastructure at scale.
  • A deep systems background — load balancing, scheduling, cluster orchestration, autoscaling, cache-coherent distributed state, high-performance networking, or similar.
  • Experience shipping performance or efficiency improvements in large-scale systems, with quantified impact including cost considerations.
  • Experience running production infrastructure with on-call, incident response, capacity events, deploy discipline.
  • A results-oriented, impact-driven approach, balancing throughput, latency, cost, stability and release velocity.
  • Ability to build strong relationships across team boundaries; a seam role.
  • Curiosity about machine learning systems and transformer inference.

Responsibilities

  • Own the technical roadmap for how the inference fleet is coordinated — where traffic goes, where capacity lives, how caches are placed, how fast the system reacts to demand, and the protocols that keep the control plane and the inference engines in sync.
  • Partner with the product, inference engine, performance, and capacity teams to identify throughput, latency, utilization, and cost wins, then turn those into shipped improvements with measurable results
  • Build the group's habit of quantitative modeling: claim a win only when you can measure it, and know before you ship what the expected effect is
  • Set technical strategy for how the control plane evolves across heterogeneous hardware, across multiple cloud providers, and across all our serving surfaces
  • Run the group's operational backbone — on-call rotations, incident response, postmortem review, deploy safety
  • Create clarity at a seam: this group sits between the API surface, the inference engines, capacity planning, and the cloud deployment teams
  • Develop and retain strong existing teams, and hire against a high technical bar
  • Coach engineers through a roadmap where priorities shift
  • Shape team structure as the scope grows: decide where the boundaries between problem areas should sit, and grow leads who can own each
  • Pick up slack when it matters. These are small teams on a critical path; sometimes the EM is the one unblocking a stuck initiative or synthesizing a design debate

Skills

Engineering management
Distributed systems
Load balancing
Scheduling
Cluster orchestration
Autoscaling
Performance optimization
Incident response
Capacity planning
Cross-team collaboration
ML systems curiosity

Education

Bachelor's degree or equivalent

Tools

Kubernetes
Cloud platforms

Job description

Anthropic is seeking an Engineering Manager to lead the inference platform team responsible for the Claude request path and fleet coordination. You will guide a group of ML platform, infrastructure, and distributed-systems engineers to ensure throughput, reliability, and efficient scaling as models, hardware, and clouds evolve.

You will own capacity planning, latency optimization, and cross-team collaboration with product, inference engine, and cloud teams, while driving architectural decisions

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Engineering Manager: Inference Infrastructure Leader
Engineering Manager: Inference Infrastructure Leader

EngineersOfAI • New York (NY), Northern (KY)

Hybrid
USD 230,000 - 360,000
Staff Software Engineer, Scalable AI Inference Systems
Staff Software Engineer, Scalable AI Inference Systems

Anthropic • San Francisco (CA)

On-site
USD 320,000 - 485,000
Staff Software Engineer, AI Inference Systems
Staff Software Engineer, AI Inference Systems

Anthropic • New York (NY)

Hybrid
USD 320,000 - 485,000
Competitive compensation
Equity donation matching (optional)
Generous vacation
+3
Engineering Manager, AI Inference Platform — Flexible Hours
Engineering Manager, AI Inference Platform — Flexible Hours

Neura Market • San Francisco (CA), Northern (KY)

Hybrid
USD 405,000 - 625,000
Engineering Manager, Inference Infrastructure
Engineering Manager, Inference Infrastructure

EngineersOfAI • New York (NY), Northern (KY)

Hybrid
USD 230,000 - 360,000
Engineering Manager, ML Infrastructure & Fleet
Engineering Manager, ML Infrastructure & Fleet

Anthropic • New York (NY)

Hybrid
USD 405,000 - 625,000
Competitive compensation
Equity donation
Vacation and parental leave
+2
Staff Cloud Inference Launch Engineer for Scalable AI
Staff Cloud Inference Launch Engineer for Scalable AI

Anthropic • San Francisco (CA)

On-site
USD 320,000 - 485,000
Staff + Senior Software Engineer, Inference Infrastructure
Staff + Senior Software Engineer, Inference Infrastructure

Menlo Ventures • San Francisco (CA)

On-site
USD 150,000 - 210,000
Staff Software Engineer: Scalable AI Inference Systems
Staff Software Engineer: Scalable AI Inference Systems

Anthropic • Seattle (WA)

Hybrid
USD 320,000 - 485,000
Lead, Inference Infrastructure & Fleet
Lead, Inference Infrastructure & Fleet

Anthropic • San Francisco (CA)

Hybrid
USD 405,000 - 625,000
Competitive compensation
Equity donation matching
Generous vacation and parental leave
+2