Customer‑Facing AI Systems Engineer – Scale & Deploy

AI Chopping Block, Inc.

San Francisco, Northern (CA, KY)

Hybrid

USD 180,000 - 230,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Competitive comp
Equity
Full benefits
Flexible PTO
Winter break
Parental leave
Fertility stipend
401(k) plan
ML startup exposure

Job summary

Baseten is seeking a Forward Deployed Engineer to work with leading AI companies, owning outcomes on Baseten across the model lifecycle. You will frame customer problems, define success criteria, build PoCs, and drive production deployments with focus on reliability and scalability.

You will collaborate with customers and internal teams, communicate clearly across engineers and leadership, and expand Baseten's infrastructure to support fast, repeatable deployments at scale.

Qualifications

  • Minimum 1-2 years of software engineering experience, shipping and maintaining code in large production systems.
  • Experience debugging complex production issues - logs, metrics, and traces to root-cause problems in unfamiliar systems.
  • Confidence owning ambiguous technical problems and making decisions under uncertainty, knowing when to pull in others.
  • Motivation beyond pure engineering: interest in working directly with customers and influencing the product.
  • Clear communication on complex technical topics with customers' engineers or leadership.
  • Genuine curiosity about AI inference and training and desire to become an infrastructure expert.
  • Willingness to respond to customers outside regular hours and participate in an on-call rotation.
  • Excitement about solving problems for large, fast-growing AI workloads.

Responsibilities

  • Act as each account's de facto CTO on Baseten, accountable for workload design, execution, and scalability.
  • Take customer objectives from vague to shipped: frame the problem, define the spec, build the PoC, and productionize quickly.
  • Design evals and benchmarks to isolate quality or performance gaps and close them via inference optimization or post-training improvements.
  • Be the first responder to mission-critical failures, triaging and owning fixes or routing to the owning team until it ships.
  • Build internal systems to make each engagement faster, including tooling for eval and deployment infra and reference implementations.
  • Shape the product by channeling account needs into the roadmap and shipping fixes/features into Baseten's codebase.
  • Handle multiple accounts simultaneously, sequencing work and keeping customers and stakeholders aligned on status and risk.

Skills

Software engineering
Production debugging
Ownership under uncertainty
Customer-facing
Communication
AI inference curiosity
On-call readiness
Problem solving at scale

Tools

Kubernetes
Slurm
Ray
GPUs
PyTorch / JAX
vLLM / TensorRT-LLM

Job description

Baseten is seeking a Forward Deployed Engineer to work with leading AI companies, owning outcomes on Baseten across the model lifecycle. You will frame customer problems, define success criteria, build PoCs, and drive production deployments with focus on reliability and scalability.

You will collaborate with customers and internal teams, communicate clearly across engineers and leadership, and expand Baseten's infrastructure to support fast, repeatable deployments at scale.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Customer-Facing AI Platform Engineer
Customer-Facing AI Platform Engineer

The Consensus • New York (NY)

On-site
USD 120,000 - 170,000
Competitive compensation
Medical, dental, vision insurance
Flexible PTO
Forward-Deployed AI Engineer for Production Scale
Forward-Deployed AI Engineer for Production Scale

The Consensus • San Francisco (CA)

On-site
USD 130,000 - 190,000
Competitive equity and compensation
100% health, dental, vision insurance
Flexible PTO including Winter Break
+4
Customer-Facing AI Platform Engineer
Customer-Facing AI Platform Engineer

Baseten • New York (NY)

On-site
USD 200,000 - 400,000
Medical, dental & vision for employee+
Flexible PTO including Winter Break
Paid parental leave
+3
Forward-Deployed AI Inference Engineer
Forward-Deployed AI Inference Engineer

BaseTen • New York (NY), San Francisco (CA)

Hybrid
USD 140,000 - 210,000
Equity
Healthcare coverage
Flexible PTO & Winter Break
+4
AI Solutions Architect: Customer Discovery & Deployments
AI Solutions Architect: Customer Discovery & Deployments

Baseten • United States

Remote
USD 140,000 - 210,000
Competitive compensation
Senior Product Manager — AI Infrastructure & Scale
Senior Product Manager — AI Infrastructure & Scale

Baseten • San Francisco (CA)

On-site
USD 210,000 - 280,000
Competitive equity
Medical, dental, vision insurance
Flexible PTO including Winter Break
+4
Remote Partner Platform Engineer for AI Integrations
Remote Partner Platform Engineer for AI Integrations

ContentBuffer • San Francisco (CA)

Remote
USD 140,000 - 190,000
Competitive compensation
Equity
Medical, dental, vision insurance
Customer-Facing AI & Deployments Lead
Customer-Facing AI & Deployments Lead

Middesk • United States

Remote
USD 140,000 - 190,000
AI Inference Engineer
AI Inference Engineer

BaseTen • New York (NY), San Francisco (CA)

On-site
USD 140,000 - 210,000
Equity
Healthcare coverage
Flexible PTO & Winter Break
+4
AI Inference Engineer
AI Inference Engineer

The Consensus • San Francisco (CA)

On-site
USD 130,000 - 190,000
Competitive equity and compensation
100% health, dental, vision insurance
Flexible PTO including Winter Break
+4