Engineering Manager, Forward Deployed AI & LLM Inference

BaseTen

New York, San Francisco (NY, CA)

On-site

USD 180,000 - 240,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive equity
Medical, dental, vision coverage
Flexible PTO
Parental leave
401(k)
Learning & networking opportunities

Job summary

Baseten is hiring an Engineering Manager (Player & Coach) to lead a team of Forward Deployed Engineers building and optimizing LLM inference workloads. You will blend hands‑on technical ownership with people leadership across design, deployment, and production monitoring.

You will partner with product and infrastructure teams to ensure high performance, reliable, and cost‑efficient AI services for Baseten customers, guiding strategic initiatives and customer engagements.

Qualifications

  • Bachelor’s, Master’s, or Ph.D. in Computer Science, Engineering, or related field.
  • 4+ years of professional software engineering experience, including 1+ year in a leadership or mentorship capacity.
  • Strong programming skills in Python, with production experience in building or optimizing ML inference systems.
  • Proven experience with LLMs, inference optimization, or serving frameworks (e.g., vLLM, TensorRT, Triton, Hugging Face, Ray Serve).
  • Familiarity with observability, profiling, and cost/performance tradeoffs in production ML systems.
  • Excellent communication and collaboration skills—able to lead cross‑functional efforts and drive outcomes in ambiguous, fast‑paced environments.

Responsibilities

  • Lead, mentor, and grow a team of Forward Deployed Engineers, providing guidance on technical direction, project execution, and professional development.
  • Set clear goals and ensure timely, high-quality delivery across multiple customer‑facing projects involving LLM deployment and inference optimization.
  • Collaborate with leadership to align team priorities with company and customer goals, balancing short‑term delivery, widely varying customer priorities, and long‑term technical initiatives.
  • Player‑coach – drive strategic product initiatives and customer engagements, combining hands‑on work with management.

Skills

Python
LLMs
Leadership
Production ML

Education

Bachelor’s/Master’s/PhD in CS or related

Tools

vLLM
TensorRT
Triton
Hugging Face
Ray Serve

Job description

Baseten is hiring an Engineering Manager (Player & Coach) to lead a team of Forward Deployed Engineers building and optimizing LLM inference workloads. You will blend hands‑on technical ownership with people leadership across design, deployment, and production monitoring.

You will partner with product and infrastructure teams to ensure high performance, reliable, and cost‑efficient AI services for Baseten customers, guiding strategic initiatives and customer engagements.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Engineering Manager, Forward-Deployed AI Infra (LLM)
Engineering Manager, Forward-Deployed AI Infra (LLM)

fde • San Francisco (CA), Northern (KY)

Hybrid
USD 260,000 - 380,000
Production AI Inference Engineer
Production AI Inference Engineer

Neura Market • San Francisco (CA)

On-site
USD 140,000 - 210,000
Equity
Premium benefits
Flexible PTO
+4
Forward-Deployed AI Inference Engineer
Forward-Deployed AI Inference Engineer

baseten • United States

On-site
USD 140,000 - 200,000
Competitive equity
Medical, dental, vision insurance for
Flexible PTO
+4
Forward-Deployed AI Inference Engineer
Forward-Deployed AI Inference Engineer

BaseTen • New York (NY), San Francisco (CA)

Hybrid
USD 140,000 - 210,000
Equity
Healthcare coverage
Flexible PTO & Winter Break
+4
Engineering Manager - Forward Deployed Engineering (LLM)
Engineering Manager - Forward Deployed Engineering (LLM)

BaseTen • New York (NY), San Francisco (CA)

On-site
USD 180,000 - 240,000
Competitive equity
Medical, dental, vision coverage
Flexible PTO
+3
Engineering Manager - Forward Deployed Engineering (LLM)
Engineering Manager - Forward Deployed Engineering (LLM)

The Consensus • New York (NY)

On-site
USD 130,000 - 160,000
Competitive compensation with equity
100% coverage of medical, dental, and vision insurance
Flexible PTO policy
Engineering Manager, Forward Deployed Engineering (LLM)
Engineering Manager, Forward Deployed Engineering (LLM)

fde • San Francisco (CA), Northern (KY)

Hybrid
USD 260,000 - 380,000
Engineering Manager, Forward Deployed AI
Engineering Manager, Forward Deployed AI

The Consensus • New York (NY)

On-site
USD 130,000 - 160,000
Competitive compensation with equity
100% coverage of medical, dental, and vision insurance
Flexible PTO policy
Forward-Deployed AI Engineer for Production Scale
Forward-Deployed AI Engineer for Production Scale

The Consensus • San Francisco (CA)

On-site
USD 130,000 - 190,000
Competitive equity and compensation
100% health, dental, vision insurance
Flexible PTO including Winter Break
+4
Forward-Deployed AI Engineer: Scale Inference & Ship Solutions
Forward-Deployed AI Engineer: Scale Inference & Ship Solutions

Triwill Group • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Equity
Medical, dental, vision coverage fore
Flexible PTO
+4