ML Performance Engineering Manager

Google

Sunnyvale (CA)

On-site

USD 207,000 - 300,000

Full time

38 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Google Sunnyvale is seeking an Engineering Manager for the Core Machine Learning group to optimize ML training and inference on TPUs. You will lead engineers across multiple teams, guiding projects from concept to production, ensuring hardware-aware software performance.

Responsibilities include driving benchmarks, collaborating with cross-functional partners, and shaping TPU-friendly models that scale for Google Cloud and internal workloads.

Qualifications

  • Bachelor’s degree or equivalent practical experience.
  • 8 years of experience in software development.
  • 5 years of experience leading ML design and optimizing ML infrastructure (e.g., model deployment, model evaluation, data processing, debugging, fine tuning).
  • 3 years of experience in a technical leadership role.
  • 2 years of experience in a people management or team leadership role.
  • Experience with ML performance analysis, benchmarking, and computer architecture.

Responsibilities

  • Lead a team of software engineers focused on identifying and maintaining ML training and serving benchmarks that are representative to Google production and the broader ML industry.
  • Achieve performance for customer launches, and in case of third-party/open-source software (OSS) models, for engaged benchmark submissions (ML Commons, InferenceX, etc.).
  • Use benchmarks to identify performance opportunities and drive both near-term SOTA (e.g., custom kernels) and out-of the box performance (compiler/runtime optimizations, agentic tooling, auto-sharding) directly and in collaboration with partner teams.
  • Participate in algorithmic innovations exploiting new TPU hardware features and model-preserving optimizations (speculative decoding, sparsity, quantization, LoRA, etc.).
  • Participate in co-designing models that are TPU-friendly to showcase model quality at performance advanced to OSS models typically designed on GPUs.

Skills

Bachelor’s degree or equivalent pr
8+ years software development
5 years ML design leadership
3 years technical leadership
2 years people management
ML performance analysis

Education

Master’s degree or PhD in Engineering/CS

Tools

CUDA
Triton
Pallas
MLIR
OpenXLA
PyTorch
JAX
vLLM

Job description

Google Sunnyvale is seeking an Engineering Manager for the Core Machine Learning group to optimize ML training and inference on TPUs. You will lead engineers across multiple teams, guiding projects from concept to production, ensuring hardware-aware software performance.

Responsibilities include driving benchmarks, collaborating with cross-functional partners, and shaping TPU-friendly models that scale for Google Cloud and internal workloads.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

ML Performance Engineering Manager
ML Performance Engineering Manager

Google • United States

On-site
USD 207,000 - 300,000
Health, dental, vision
401(k) with company match
Paid time off 20 days
+4
TPU Performance Engineer – ML Hardware & Compiler Co-Design
TPU Performance Engineer – ML Hardware & Compiler Co-Design

Google • Sunnyvale (CA)

On-site
USD 150,000 - 210,000
Bonus target
Equity
Benefits
Engineering Manager, ML Performance
Engineering Manager, ML Performance

Google • Sunnyvale (CA)

On-site
USD 207,000 - 300,000
ML Hardware Architect — TPU AI Accelerators
ML Hardware Architect — TPU AI Accelerators

Google Inc. • Sunnyvale (CA)

On-site
USD 163,000 - 236,000
Staff Software Engineer, TPU Performance & ML Infrastructure
Staff Software Engineer, TPU Performance & ML Infrastructure

Google • New York (NY)

On-site
USD 207,000 - 300,000
Senior Staff TPU Performance Co-Design Engineer
Senior Staff TPU Performance Co-Design Engineer

Google • Town of Montana (WI)

On-site
USD 240,000 - 333,000
Equity
Benefits
Staff Software Engineer - TPU Performance & ML Infra
Staff Software Engineer - TPU Performance & ML Infra

Google • Ionia (NY)

On-site
USD 207,000 - 300,000
Staff ML Systems Architect - Co-Design & TPU Performance
Staff ML Systems Architect - Co-Design & TPU Performance

Google • Mountain View (CA)

On-site
USD 207,000 - 300,000
Senior TPU Software Engineer for ML Performance
Senior TPU Software Engineer for ML Performance

Google • Sunnyvale (CA)

On-site
USD 207,000 - 300,000
Equity
Bonus target 20%
Company benefits
Staff ML Systems Co-Design Engineer (Equity)
Staff ML Systems Co-Design Engineer (Equity)

Google Inc. • Mountain View (CA)

On-site
USD 207,000 - 300,000