Staff ML Engineer - Simulation & Efficient Inference

Waymo

United States

On-site

USD 251,000 - 310,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Waymo is seeking a seasoned ML systems engineer to advance the Waymo Driver’s AI stack. You will analyze architectures, reduce training and inference bottlenecks, and implement efficient techniques like quantization and distillation.

The role emphasizes cross‑device optimization on TPUs/GPUs and scalable data/model parallelism. The ideal candidate holds a MS/PhD in a relevant field with 5+ years of deep learning experience, strong Python and C++ skills, and hands-on work with JAX, Flax, and ML

Qualifications

  • MS or PhD in Computer Science, Machine Learning, Robotics, or related field.
  • 5+ years of experience with deep learning architectures, optimization techniques, and ML systems.
  • Proficiency in Python; C++ a plus; experience with ML frameworks.

Responsibilities

  • Analyze model architectures and identify bottlenecks in training and inference performance.
  • Apply and develop techniques such as quantization, pruning, distillation, and efficient attention mechanisms.
  • Optimize model code for hardware accelerators (TPUs, GPUs) using compiler features and low-level libraries (XLA).
  • Experiment with model partitioning and sharding strategies to improve scalability and efficiency.
  • Design and implement low-latency serving solutions and optimize training pipelines to reduce time.
  • Build and maintain tools for performance analysis, profiling, and debugging of ML models.

Skills

Transformers MoEs
Diffusion Models
Profiling tools
Python
C++
XLA
ML optimization

Education

MS or PhD in CS/ML/Robotics

Tools

JAX
Flax
TensorFlow/PyTorch

Job description

Waymo is seeking a seasoned ML systems engineer to advance the Waymo Driver’s AI stack. You will analyze architectures, reduce training and inference bottlenecks, and implement efficient techniques like quantization and distillation.

The role emphasizes cross‑device optimization on TPUs/GPUs and scalable data/model parallelism. The ideal candidate holds a MS/PhD in a relevant field with 5+ years of deep learning experience, strong Python and C++ skills, and hands-on work with JAX, Flax, and ML

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior ML Engineer - Simulation Realism & Data Quality
Senior ML Engineer - Simulation Realism & Data Quality

Neura Market • Mountain View (CA)

Hybrid
USD 251,000 - 310,000
Equity incentive plan
Discretionary annual bonus program
Generous benefits program
Senior ML Engineer, Perception & Real-World Systems
Senior ML Engineer, Perception & Real-World Systems

Neura Market • Mountain View (CA)

Hybrid
USD 251,000 - 310,000
Staff ML Engineer, Generative Model Performance & Efficiency
Staff ML Engineer, Generative Model Performance & Efficiency

Waymo • Mountain View (NY)

On-site
USD 251,000 - 310,000
Discretionary annual bonus program
Equity incentive plan
Generous company benefits
ML Engineer - Real-Time Planning & Selection
ML Engineer - Real-Time Planning & Selection

Waymo • Mountain View (CA)

On-site
USD 175,000 - 215,000
Discretionary annual bonus
Equity incentive plan
Staff ML Engineer, Generative Model Performance & Efficiency
Staff ML Engineer, Generative Model Performance & Efficiency

Waymo • Mountain View (CA)

On-site
USD 251,000 - 310,000
Annual bonus program
Equity incentive plan
Generous Company benefits program
Staff ML Engineer: Gen AI & RL for Autonomous Driving
Staff ML Engineer: Gen AI & RL for Autonomous Driving

Neura Market • Mountain View (CA)

Hybrid
USD 238,000 - 302,000
Senior ML Systems Engineer - RL at Scale (Hybrid)
Senior ML Systems Engineer - RL at Scale (Hybrid)

Neura Market • Mountain View (CA)

Hybrid
USD 204,000 - 259,000
Discretionary annual bonus
Equity incentive plan
Benefits program
Staff ML Engineer: Ultra-Fast Model Serving & Optimization
Staff ML Engineer: Ultra-Fast Model Serving & Optimization

Waymo • Mountain View (NY)

On-site
USD 251,000 - 310,000
Discretionary annual bonus program
Equity incentive plan
Generous company benefits
Staff ML Engineer: Evaluation & GenAI for Autonomous Driving
Staff ML Engineer: Evaluation & GenAI for Autonomous Driving

Neura Market • Mountain View (CA)

Hybrid
USD 238,000 - 302,000
Discretionary annual bonus
Equity incentive plan
Benefits program
ML Planner Engineer - Real-Time Autonomous Driving
ML Planner Engineer - Real-Time Autonomous Driving

Waymo • San Francisco (CA)

On-site
USD 175,000 - 215,000
Equity incentives
Discretionary annual bonus