ML Systems Engineer: TPU Efficiency & Optimization

Google LLC

Mountain View (CA)

On-site

USD 147,000 - 210,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Google's YouTube team in Mountain View seeks a Software Engineer III, ML to optimize TPU-based model efficiency. You will focus on reducing training and serving costs through techniques like low-precision quantization, knowledge distillation, and hardware-friendly model design.

You'll build and optimize models in the Recommendation System stack and collaborate across ML and hardware teams to scale TPU workloads and improve overall pipeline performance.

Qualifications

  • Bachelor's degree or equivalent practical experience.
  • 2 years of experience programming in C++ or Python.
  • 2 years of experience with software design and architecture.
  • 2 years of experience testing, and launching software products.
  • Experience with ML model optimization.
  • Experience with ML frameworks such as TensorFlow, JAX, and PyTorch, or ML compilers (e.g., XLA).

Responsibilities

  • Profile ML workloads and optimize accelerator utilization.
  • Explore and productionize efficiency techniques like quantization and distillation.
  • Collaborate with ML model developers and hardware teams to design efficient architectures.
  • Support real-time training and inference in the YouTube Recommendation stack.

Skills

C++/Python programming
Software design
Testing & deployment
ML optimization
ML frameworks

Education

Bachelor's degree or equivalent
Master's degree or PhD in Computer Science

Tools

XLA
TPU/GPUs familiarity

Job description

Google's YouTube team in Mountain View seeks a Software Engineer III, ML to optimize TPU-based model efficiency. You will focus on reducing training and serving costs through techniques like low-precision quantization, knowledge distillation, and hardware-friendly model design.

You'll build and optimize models in the Recommendation System stack and collaborate across ML and hardware teams to scale TPU workloads and improve overall pipeline performance.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

ML TPU Efficiency Engineer
ML TPU Efficiency Engineer

Google • Mountain View (CA)

On-site
USD 147,000 - 210,000
15% bonus target
Equity
Benefits
Software Engineer III, ML, TPU Efficiency, YouTube
Software Engineer III, ML, TPU Efficiency, YouTube

Google • Mountain View (CA)

On-site
USD 147,000 - 210,000
15% bonus target
Equity
Benefits
Staff ML Systems Architect - Co-Design & TPU Performance
Staff ML Systems Architect - Co-Design & TPU Performance

Google • Mountain View (CA)

On-site
USD 207,000 - 300,000
ML Performance Engineering Manager
ML Performance Engineering Manager

Google • Sunnyvale (CA)

On-site
USD 207,000 - 300,000
Staff ML Systems Co-Design Engineer (Equity)
Staff ML Systems Co-Design Engineer (Equity)

Google Inc. • Mountain View (CA)

On-site
USD 207,000 - 300,000
Senior TPU Software Engineer for ML Performance
Senior TPU Software Engineer for ML Performance

Google • Sunnyvale (CA)

On-site
USD 207,000 - 300,000
Equity
Bonus target 20%
Company benefits
Staff Software Engineer, TPU Performance & ML Infrastructure
Staff Software Engineer, TPU Performance & ML Infrastructure

Google • New York (NY)

On-site
USD 207,000 - 300,000
TPU Performance Engineer – ML Hardware & Compiler Co-Design
TPU Performance Engineer – ML Hardware & Compiler Co-Design

Google • Sunnyvale (CA)

On-site
USD 150,000 - 210,000
Bonus target
Equity
Benefits
ML Engineer: Edge AI & Model Optimization
ML Engineer: Edge AI & Model Optimization

Google • Des Moines (IA)

Remote
TPU Architecture Modeling Engineer
TPU Architecture Modeling Engineer

Google Inc. • Sunnyvale (CA)

On-site
USD 138,000 - 197,000
Equity grants
Benefits