Forward-Deployed AI Inference Engineer

BaseTen

New York, San Francisco (NY, CA)

Hybrid

USD 140,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Healthcare coverage
Flexible PTO & Winter Break
Parental leave
Fertility stipend
401(k)
Learning opportunities

Job summary

Baseten is seeking a Forward Deployed Engineer to partner with customers, architecting, building, and deploying high-scale AI applications on Baseten's platform. You will own the customer journey from exploration to production deployment, translating business goals into observable services with clear quality, latency, and cost outcomes.

This role blends engineering, product management, and customer success, with hands-on coding and cross-functional collaboration across product and engineering

Qualifications

  • Bachelor's, Master's, or Ph.D. in Computer Science, Engineering, Mathematics, or related field.
  • 2+ years of professional work experience in a fast-paced, high-growth environment.
  • Demonstrated experience with one or more general-purpose programming languages, with a strong preference for Python.
  • Familiarity with AI/ML pipelines and the lifecycle of ML model development and deployment.
  • Strong communication skills, particularly on complex technical topics.
  • Experience in building or optimizing AI/ML projects is highly valued.

Responsibilities

  • Develop and maintain software systems and product features in a production-level environment, with a preference for Python.
  • Drive customer impact by designing, implementing, and deploying Baseten solutions end-to-end across the customer journey.
  • Deliver with velocity by turning vague objectives into clear specs and PoCs.
  • Optimize and enhance AI/ML projects, contributing to the technical stack.
  • Own products and customer projects end-to-end as engineer, PM, and customer-facing lead.
  • Navigate ambiguity and make informed tradeoffs to solve problems.
  • Demonstrate pride, ownership, and accountability for your work.

Skills

Python
General-purpose languages
Communication skills
AI/ML workflows

Education

Bachelor's/Master's/Ph.D. in CS/Engineering/Math or related

Job description

Baseten is seeking a Forward Deployed Engineer to partner with customers, architecting, building, and deploying high-scale AI applications on Baseten's platform. You will own the customer journey from exploration to production deployment, translating business goals into observable services with clear quality, latency, and cost outcomes.

This role blends engineering, product management, and customer success, with hands-on coding and cross-functional collaboration across product and engineering

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Forward-Deployed AI Inference Engineer
Forward-Deployed AI Inference Engineer

baseten • United States

On-site
USD 140,000 - 200,000
Competitive equity
Medical, dental, vision insurance for
Flexible PTO
+4
Production AI Inference Engineer
Production AI Inference Engineer

Neura Market • San Francisco (CA)

On-site
USD 140,000 - 210,000
Equity
Premium benefits
Flexible PTO
+4
Forward-Deployed AI Engineer for Production Scale
Forward-Deployed AI Engineer for Production Scale

The Consensus • San Francisco (CA)

On-site
USD 130,000 - 190,000
Competitive equity and compensation
100% health, dental, vision insurance
Flexible PTO including Winter Break
+4
Forward-Deployed AI Engineer for Post-Training Deployments
Forward-Deployed AI Engineer for Post-Training Deployments

Neura Market • San Francisco (CA)

On-site
USD 150,000 - 230,000
Competitive equity-based compensation
100% health insurance for employee and
dependents
+6
Customer-Facing AI Platform Engineer
Customer-Facing AI Platform Engineer

The Consensus • New York (NY)

On-site
USD 120,000 - 170,000
Competitive compensation
Medical, dental, vision insurance
Flexible PTO
Forward-Deployed AI Engineer: Scale Inference & Ship Solutions
Forward-Deployed AI Engineer: Scale Inference & Ship Solutions

Triwill Group • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Equity
Medical, dental, vision coverage fore
Flexible PTO
+4
Customer-Facing AI Platform Engineer
Customer-Facing AI Platform Engineer

Baseten • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 260,000
Equity
Medical, dental, vision coverage for +
Flexible PTO including Winter Break
+4
Customer‑Facing AI Systems Engineer – Scale & Deploy
Customer‑Facing AI Systems Engineer – Scale & Deploy

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 230,000
Competitive comp
Equity
Full benefits
+6
Engineering Manager, Forward Deployed AI & LLM Inference
Engineering Manager, Forward Deployed AI & LLM Inference

BaseTen • New York (NY), San Francisco (CA)

On-site
USD 180,000 - 240,000
Competitive equity
Medical, dental, vision coverage
Flexible PTO
+3
AI Inference Engineer
AI Inference Engineer

baseten • United States

On-site
USD 140,000 - 200,000
Competitive equity
Medical, dental, vision insurance for
Flexible PTO
+4