Engineering Manager, AI Inference Platform — Flexible Hours

Neura Market

San Francisco, Northern (CA, KY)

Hybrid

USD 405,000 - 625,000

Full time

7 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Anthropic, based in San Francisco, is seeking an engineering leader to own the technical roadmap for coordinating the inference fleet and to guide a diverse group of ML platform, infrastructure, and distributed-systems engineers. You’ll shape architecture, drive reliability, and ensure the end-to-end path from request to model stays efficient as models and clouds evolve.

You will partner with product, inference engine, performance, and capacity teams to push throughput, latency, and cost

Qualifications

  • Engineering management experience leading critical-path production infrastructure at scale.
  • Deep systems background—load balancing, scheduling, autoscaling, distributed state.
  • Experience shipping performance improvements with measurable impact, including cost.
  • Experience running production infrastructure with on-call, incident response.
  • Results-oriented with balance of throughput, latency, cost, and delivery timelines.
  • Ability to build strong relationships across team boundaries ( seam role ).
  • Curiosity about machine learning systems and transformer inference.

Responsibilities

  • Own the technical roadmap for how the inference fleet is coordinated: traffic routing, capacity placement, and synchronization protocols.
  • Partner with product, inference engine, performance, and capacity teams to identify throughput, latency, utilization, and cost wins and ship improvements.
  • Build the group’s habit of quantitative modeling and measure impact before shipping.
  • Set technical strategy for control plane evolution across heterogeneous hardware and multiple clouds.
  • Run the group’s on-call rotations, incident response, deploy safety, and ensure stability while shipping aggressively.
  • Create clarity at the seam between API surface, inference engines, capacity planning, and cloud deployment.

Skills

Engineering management
Systems depth
Performance optimization
Production infrastructure
Cross-team leadership
ML systems curiosity

Education

Bachelor's degree

Tools

Kubernetes

Job description

Anthropic, based in San Francisco, is seeking an engineering leader to own the technical roadmap for coordinating the inference fleet and to guide a diverse group of ML platform, infrastructure, and distributed-systems engineers. You’ll shape architecture, drive reliability, and ensure the end-to-end path from request to model stays efficient as models and clouds evolve.

You will partner with product, inference engine, performance, and capacity teams to push throughput, latency, and cost

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Engineering Manager: Inference Infrastructure Leader
Engineering Manager: Inference Infrastructure Leader

EngineersOfAI • New York (NY), Northern (KY)

Hybrid
USD 230,000 - 360,000
Lead, Inference Infrastructure & Fleet
Lead, Inference Infrastructure & Fleet

Anthropic • San Francisco (CA)

Hybrid
USD 405,000 - 625,000
Competitive compensation
Equity donation matching
Generous vacation and parental leave
+2
Engineering Manager, Inference Infrastructure
Engineering Manager, Inference Infrastructure

EngineersOfAI • New York (NY), Northern (KY)

Hybrid
USD 230,000 - 360,000
Staff Software Engineer, Scalable AI Inference Systems
Staff Software Engineer, Scalable AI Inference Systems

Anthropic • San Francisco (CA)

On-site
USD 320,000 - 485,000
Senior AI Inference Deployment Engineer
Senior AI Inference Deployment Engineer

Anthropic • San Francisco (CA)

On-site
USD 320,000 - 485,000
Competitive compensation
Flexible working hours
Generous vacation and parental leave
+1
Staff Software Engineer, AI Inference Systems
Staff Software Engineer, AI Inference Systems

Anthropic • New York (NY)

Hybrid
USD 320,000 - 485,000
Competitive compensation
Equity donation matching (optional)
Generous vacation
+3
Engineering Manager, Inference Infrastructure
Engineering Manager, Inference Infrastructure

Anthropic • San Francisco (CA)

Hybrid
USD 405,000 - 625,000
Competitive compensation
Equity donation matching
Generous vacation and parental leave
+2
Engineering Manager, Inference
Engineering Manager, Inference

Anthropic • San Francisco (CA)

Hybrid
USD 425,000 - 560,000
Competitive compensation
Flexible working hours
Generous vacation and parental leave
Performance Engineer — AI Inference Systems
Performance Engineer — AI Inference Systems

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Visa sponsorship
Flexible hybrid work policy
Senior AI Inference Deployment Engineer
Senior AI Inference Deployment Engineer

Anthropic • Seattle (WA)

Hybrid
USD 320,000 - 485,000
Competitive compensation
Flexible working hours
Generous vacation and parental leave