Customer-Facing AI Platform Engineer

Baseten

San Francisco, Northern (CA, KY)

Hybrid

USD 180,000 - 260,000

Full time

9 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Medical, dental, vision coverage for +
Flexible PTO including Winter Break
Paid parental leave
Fertility and family-building stipend
401(k)
Learning and networking opportunities

Job summary

Baseten is hiring Forward Deployed Engineers to work with the world’s fastest-growing AI companies, taking technical ownership for workloads from inference to deployment. You’ll collaborate across the model lifecycle, optimize performance, and drive customer success at scale.

You’ll engage with Kubernetes, Slurm, Ray, and modern inference stacks, bringing deep infrastructure expertise and a passion for solving complex AI workloads on production systems.

Qualifications

  • 1–2 years of software eng experience shipping and maintaining production code.

Responsibilities

  • Act as account CTO on Baseten, accountable for workloads’ design, run, and scale.
  • Frame customer problems, define specs, build PoC, and ship to production quickly.
  • Design evals/benchmarks to isolate quality or performance gaps and close them.
  • Be first responder to mission-critical failures, triage and fix or route to owning team.
  • Build internal tooling and automation for eval/deploy infrastructure and self-serve recipes.
  • Shape the product roadmap by channeling accounts’ needs into Baseten's codebase.
  • Manage multiple accounts, align stakeholders, and communicate status and risk.

Skills

Software engineering
Production debugging
Problem ownership
Customer focus
Technical communication
AI inference
On-call readiness
Scale problem solving

Tools

Kubernetes
Slurm
Ray
vLLM
TensorRT-LLM
SGLang
PyTorch
JAX

Job description

Baseten is hiring Forward Deployed Engineers to work with the world’s fastest-growing AI companies, taking technical ownership for workloads from inference to deployment. You’ll collaborate across the model lifecycle, optimize performance, and drive customer success at scale.

You’ll engage with Kubernetes, Slurm, Ray, and modern inference stacks, bringing deep infrastructure expertise and a passion for solving complex AI workloads on production systems.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Customer-Facing AI Platform Engineer
Customer-Facing AI Platform Engineer

The Consensus • New York (NY)

On-site
USD 120,000 - 170,000
Competitive compensation
Medical, dental, vision insurance
Flexible PTO
Forward-Deployed AI Inference Engineer
Forward-Deployed AI Inference Engineer

BaseTen • New York (NY), San Francisco (CA)

Hybrid
USD 140,000 - 210,000
Equity
Healthcare coverage
Flexible PTO & Winter Break
+4
Customer‑Facing AI Systems Engineer – Scale & Deploy
Customer‑Facing AI Systems Engineer – Scale & Deploy

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 230,000
Competitive comp
Equity
Full benefits
+6
Forward-Deployed AI Inference Engineer
Forward-Deployed AI Inference Engineer

baseten • United States

On-site
USD 140,000 - 200,000
Competitive equity
Medical, dental, vision insurance for
Flexible PTO
+4
Forward-Deployed AI Engineer for Production Scale
Forward-Deployed AI Engineer for Production Scale

The Consensus • San Francisco (CA)

On-site
USD 130,000 - 190,000
Competitive equity and compensation
100% health, dental, vision insurance
Flexible PTO including Winter Break
+4
Forward-Deployed AI Engineer: Scale Inference & Ship Solutions
Forward-Deployed AI Engineer: Scale Inference & Ship Solutions

Triwill Group • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Equity
Medical, dental, vision coverage fore
Flexible PTO
+4
Forward-Deployed AI Engineer for Post-Training Deployments
Forward-Deployed AI Engineer for Post-Training Deployments

Neura Market • San Francisco (CA)

On-site
USD 150,000 - 230,000
Competitive equity-based compensation
100% health insurance for employee and
dependents
+6
Customer-Facing AI Deployment Engineer
Customer-Facing AI Deployment Engineer

Applied Compute • San Francisco (CA)

On-site
USD 180,000 - 240,000
Competitive compensation
Equity
Generous health benefits
+5
AI Inference Engineer
AI Inference Engineer

BaseTen • New York (NY), San Francisco (CA)

Hybrid
USD 140,000 - 210,000
Equity
Healthcare coverage
Flexible PTO & Winter Break
+4
AI Inference Engineer
AI Inference Engineer

The Consensus • San Francisco (CA)

On-site
USD 130,000 - 190,000
Competitive equity and compensation
100% health, dental, vision insurance
Flexible PTO including Winter Break
+4