Platform Engineer - Scalable ML Inference

Menlo Ventures

San Francisco (CA)

On-site

USD 180,000 - 260,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Competitive compensation
Ownership culture
World-class team

Job summary

Chai Discovery is building a design suite for molecules and AI-driven biology. Platform engineers own the serving stack that makes frontier models fast, reliable, and scalable across a multi-cloud GPU fleet.

You'll optimize latency, throughput, and GPU efficiency, and turn models into product-ready pipelines used by researchers. You'll work with researchers, product engineers, and the commercial team to ship systems that researchers depend on and to drive observability, incident

Qualifications

  • 4+ years building production systems with depth in performance, distributed systems, or ML serving.
  • Experience optimizing model inference: GPU utilization, batching, quantization, caching, or kernel-level work.
  • A platform mindset: you like building the tools and abstractions that make other engineers and researchers faster.
  • End-to-end ownership of 24/7 systems, including observability, alerting, and incident response.
  • Experience across both 0-to-1 buildouts and 1-to-n scale-ups, with an always-evolving playbook you bring wherever you go.
  • The instinct to treat cost and efficiency as first-class constraints, not afterthoughts.

Responsibilities

  • Own the serving stack that turns frontier models into a product researchers depend on: latency, throughput, GPU efficiency, batching, and autoscaling across a large multi-cloud GPU fleet.
  • Contribute to pipelines and observability tooling that lets researchers ship faster.
  • Work closely with researchers, product engineers, and the commercial team deploying them to the world's largest pharma companies.
  • Drive reliability and performance in production systems with 24/7 uptime.
  • Collaborate to continuously evolve the platform with a focus on cost efficiency.

Skills

Production systems
GPU optimization
Platform mindset
Observability & incident response
End-to-end ownership
Cost optimization

Job description

Chai Discovery is building a design suite for molecules and AI-driven biology. Platform engineers own the serving stack that makes frontier models fast, reliable, and scalable across a multi-cloud GPU fleet.

You'll optimize latency, throughput, and GPU efficiency, and turn models into product-ready pipelines used by researchers. You'll work with researchers, product engineers, and the commercial team to ship systems that researchers depend on and to drive observability, incident

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer, Platform & Inference
Software Engineer, Platform & Inference

Menlo Ventures • San Francisco (CA)

On-site
USD 180,000 - 260,000
Competitive compensation
Ownership culture
World-class team
Senior ML Platform Engineer - Scale ML Infra & Pipelines
Senior ML Platform Engineer - Scale ML Infra & Pipelines

Menlo Ventures • San Francisco (CA)

Hybrid
USD 187,000 - 259,000
4 days in office per week
Fridays from home
Medical, dental, vision benefits
+3
ML Platform Engineer for AI-Driven Drug Discovery
ML Platform Engineer for AI-Driven Drug Discovery

Genentech • San Francisco (CA)

On-site
USD 147,600 - 274,000
ML Platform Engineer: Scale AI & Inference
ML Platform Engineer: Scale AI & Inference

Apply • San Francisco (CA)

Hybrid
USD 245,000 - 345,000
Flexible Time Off
Health Insurance
Work From Home Allowance
+2
ML Platform Engineer: Scale AI Infra, Deploy & Optimize
ML Platform Engineer: Scale AI Infra, Deploy & Optimize

United States Digital Space LLC • United States

Remote
USD 120,000 - 180,000
Lead ML Platform Engineer: Training & Inference at Scale
Lead ML Platform Engineer: Training & Inference at Scale

Paramount • Burbank (CA)

On-site
USD 157,000 - 235,000
Benefits package
On-site & virtual events
Generous PTO
Principal ML Platform Engineer - Scalable, Reliable Systems
Principal ML Platform Engineer - Scalable, Reliable Systems

EngineersOfAI • Sunnyvale (CA)

On-site
USD 150,000 - 200,000
ML Platform Engineer — Scale GPU-Driven Research Infra
ML Platform Engineer — Scale GPU-Driven Research Infra

Neura Market • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
ML Platform Infra Engineer for Scalable AI Serving
ML Platform Infra Engineer for Scalable AI Serving

Snap Inc. • Los Angeles (CA)

On-site
USD 157,000 - 235,000
Paid parental leave
Comprehensive medical coverage
Mental health support programs
+1
Senior Platform Engineer, Inference & GPU Compute Infra
Senior Platform Engineer, Inference & GPU Compute Infra

Together • San Francisco (CA)

On-site
USD 240,000 - 280,000
Startup equity
Health insurance