Customer‑Facing AI Systems Engineer – Scale & Deploy

AI Chopping Block

San Francisco, Northern (CA, KY)

Hybrid

USD 180,000 - 230,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive comp
Equity
Full benefits
Flexible PTO
Winter break
Parental leave
Fertility stipend
401(k) plan
ML startup exposure

Job summary

Baseten is seeking a Forward Deployed Engineer to work with leading AI companies, owning outcomes on Baseten across the model lifecycle. You will frame customer problems, define success criteria, build PoCs, and drive production deployments with focus on reliability and scalability.

You will collaborate with customers and internal teams, communicate clearly across engineers and leadership, and expand Baseten's infrastructure to support fast, repeatable deployments at scale.

Qualifications

  • Minimum 1-2 years of software engineering experience, shipping and maintaining code in large production systems.
  • Experience debugging complex production issues - logs, metrics, and traces to root-cause problems in unfamiliar systems.
  • Confidence owning ambiguous technical problems and making decisions under uncertainty, knowing when to pull in others.
  • Motivation beyond pure engineering: interest in working directly with customers and influencing the product.
  • Clear communication on complex technical topics with customers' engineers or leadership.
  • Genuine curiosity about AI inference and training and desire to become an infrastructure expert.
  • Willingness to respond to customers outside regular hours and participate in an on-call rotation.
  • Excitement about solving problems for large, fast-growing AI workloads.

Responsibilities

  • Act as each account's de facto CTO on Baseten, accountable for workload design, execution, and scalability.
  • Take customer objectives from vague to shipped: frame the problem, define the spec, build the PoC, and productionize quickly.
  • Design evals and benchmarks to isolate quality or performance gaps and close them via inference optimization or post-training improvements.
  • Be the first responder to mission-critical failures, triaging and owning fixes or routing to the owning team until it ships.
  • Build internal systems to make each engagement faster, including tooling for eval and deployment infra and reference implementations.
  • Shape the product by channeling account needs into the roadmap and shipping fixes/features into Baseten's codebase.
  • Handle multiple accounts simultaneously, sequencing work and keeping customers and stakeholders aligned on status and risk.

Skills

Software engineering
Production debugging
Ownership under uncertainty
Customer-facing
Communication
AI inference curiosity
On-call readiness
Problem solving at scale

Tools

Kubernetes
Slurm
Ray
GPUs
PyTorch / JAX
vLLM / TensorRT-LLM

Job description

Baseten is seeking a Forward Deployed Engineer to work with leading AI companies, owning outcomes on Baseten across the model lifecycle. You will frame customer problems, define success criteria, build PoCs, and drive production deployments with focus on reliability and scalability.

You will collaborate with customers and internal teams, communicate clearly across engineers and leadership, and expand Baseten's infrastructure to support fast, repeatable deployments at scale.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Customer-Facing AI Platform Engineer
Customer-Facing AI Platform Engineer

The Consensus • New York (NY)

On-site
USD 120,000 - 170,000
Competitive compensation
Medical, dental, vision insurance
Flexible PTO
Forward-Deployed AI Engineer for Production Scale
Forward-Deployed AI Engineer for Production Scale

The Consensus • San Francisco (CA)

On-site
USD 130,000 - 190,000
Competitive equity and compensation
100% health, dental, vision insurance
Flexible PTO including Winter Break
+4
Forward-Deployed AI Engineer for Post-Training Deployments
Forward-Deployed AI Engineer for Post-Training Deployments

Neura Market • San Francisco (CA)

On-site
USD 150,000 - 230,000
Competitive equity-based compensation
100% health insurance for employee and
dependents
+6
Customer-Facing AI Platform Engineer
Customer-Facing AI Platform Engineer

Baseten • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 260,000
Equity
Medical, dental, vision coverage for +
Flexible PTO including Winter Break
+4
Forward-Deployed AI Inference Engineer
Forward-Deployed AI Inference Engineer

BaseTen • New York (NY), San Francisco (CA)

Hybrid
USD 140,000 - 210,000
Equity
Healthcare coverage
Flexible PTO & Winter Break
+4
Forward-Deployed AI Inference Engineer
Forward-Deployed AI Inference Engineer

baseten • United States

On-site
USD 140,000 - 200,000
Competitive equity
Medical, dental, vision insurance for
Flexible PTO
+4
Production AI Inference Engineer
Production AI Inference Engineer

Neura Market • San Francisco (CA)

On-site
USD 140,000 - 210,000
Equity
Premium benefits
Flexible PTO
+4
Forward-Deployed AI Engineer: Scale Inference & Ship Solutions
Forward-Deployed AI Engineer: Scale Inference & Ship Solutions

Triwill Group • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Equity
Medical, dental, vision coverage fore
Flexible PTO
+4
Senior Product Manager — AI Infrastructure & Scale
Senior Product Manager — AI Infrastructure & Scale

Baseten • San Francisco (CA)

On-site
USD 210,000 - 280,000
Competitive equity
Medical, dental, vision insurance
Flexible PTO including Winter Break
+4
Customer-Facing AI Deployment Engineer
Customer-Facing AI Deployment Engineer

People Culture Talent • United States

On-site
USD 89,000 - 188,000