Senior ML Runtime Engineer — Real-Time Inference

Waymo

California (MO)

Hybrid

USD 213,000 - 263,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Bonus program
Equity incentive plan
Benefits

Job summary

Waymo is seeking an engineer to build the next generation ML onboard inference runtime and related deployment tooling, spanning from efficient deep learning models to optimized serving. You will work across the ML stack for both onboard compute and large-scale data center environments.

The role combines systems engineering with ML compiler and runtime work, requiring collaboration with perception, planner and research teams, and offers opportunities to influence performance across Waymo's

Qualifications

  • BS/MS in CS/EE/Deep Learning or related field.
  • 5+ years professional software engineering in ML systems.
  • 5+ years production C++ experience.
  • 3+ years Python and major deep learning frameworks (PyTorch, JAX).
  • Experience optimizing ML software for GPUs/TPUs and hardware accelerators.
  • Experience building low-latency, highly concurrent distributed backends.

Responsibilities

  • Architect and develop efficient ML runtime and serving system for onboard and data-center environments.
  • Lead integration of ML inference runtimes across onboard and offboard domains with real-time constraints.
  • Drive migration toward a JAX-native runtime (OpenXLA/PjRT, TensorRT).
  • Collaborate with Waymo ML teams to analyze workloads and apply hardware-aware optimizations.
  • Design tooling for profiling, benchmarking, and identifying end-to-end bottlenecks.

Skills

C++
Python
Deep Learning
Distributed systems

Education

B.S./M.S. in CS/EE/Deep Learning

Tools

OpenXLA/PjRT
TensorRT
ONNX Runtime
TVM
CUDA
JAX
PyTorch

Job description

Waymo is seeking an engineer to build the next generation ML onboard inference runtime and related deployment tooling, spanning from efficient deep learning models to optimized serving. You will work across the ML stack for both onboard compute and large-scale data center environments.

The role combines systems engineering with ML compiler and runtime work, requiring collaboration with perception, planner and research teams, and offers opportunities to influence performance across Waymo's

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior ML Platform Engineer — Scalable Inference & Infra
Senior ML Platform Engineer — Scalable Inference & Infra

Waymo • California (MO)

Hybrid
USD 213,000 - 263,000
Discretionary annual bonus
Equity incentive plan
Company benefits program
Staff ML Engineer - Simulation & Efficient Inference
Staff ML Engineer - Simulation & Efficient Inference

Waymo • United States

On-site
USD 251,000 - 310,000
Senior Machine Learning Engineer, Runtime and Serving
Senior Machine Learning Engineer, Runtime and Serving

Waymo • California (MO)

Hybrid
USD 213,000 - 263,000
Bonus program
Equity incentive plan
Benefits
ML Systems Engineer: Runtime & Optimization
ML Systems Engineer: Runtime & Optimization

Waymo • California (MO)

Hybrid
USD 213,000 - 263,000
Senior ML Infra Engineer - Ultra-Realistic Simulation
Senior ML Infra Engineer - Ultra-Realistic Simulation

Waymo • San Francisco (CA)

On-site
USD 213,000 - 263,000
Discretionary annual bonus program
Equity incentive plan
Generous company benefits program
Senior ML Compiler Engineer - High-Performance AI, Hybrid
Senior ML Compiler Engineer - High-Performance AI, Hybrid

Waymo • Mountain View (CA)

Hybrid
USD 213,000 - 263,000
Discretionary annual bonus
Equity incentive plan
Generous benefits program
Senior ML Architect - On-Device LLM/VLM Systems
Senior ML Architect - On-Device LLM/VLM Systems

Waymo • California (MO)

On-site
USD 298,000 - 368,000
Annual bonus
Equity incentive
Benefits program
Senior ML Data Infra Engineer — Remote/Hybrid + Equity
Senior ML Data Infra Engineer — Remote/Hybrid + Equity

Waymo • California (MO)

Hybrid
USD 213,000 - 263,000
Discretionary bonus
Equity incentive plan
Company benefits
Senior ML Engineer, Perception & Real-World Systems
Senior ML Engineer, Perception & Real-World Systems

Neura Market • Mountain View (CA)

Hybrid
USD 251,000 - 310,000
Technical Lead Manager, Scalable ML Runtime — Edge & Cloud
Technical Lead Manager, Scalable ML Runtime — Edge & Cloud

Waymo • Mountain View (CA)

On-site
USD 251,000 - 310,000