Lead ML Performance Engineering Manager

Google

Kirkland (WA)

On-site

USD 207,000 - 300,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
401(k) with company match
Paid Time Off: 20 days per year
Sick time: 40 hours/year (69 inSeattle
Maternity leave 28-30 weeks
Baby Bonding Leave 18 weeks
Holidays 13 days

Job summary

Google in Kirkland, WA is seeking an Engineering Manager for the Core ML/TPU Performance group. You will lead engineers across teams to optimize ML training and inference on TPU hardware, aligning software and hardware strategies with product goals and external collaborations.

The role emphasizes technical leadership, project budgeting, and multi-site coordination to deliver high-performance AI capabilities at scale. Applicants should have deep ML systems experience and people management skills.

Qualifications

  • Bachelor’s degree or equivalent practical experience.
  • 8 years of experience in software development.
  • 5 years of experience leading ML design and optimizing ML infrastructure (e.g., model deployment, model evaluation, data processing, debugging, fine tuning).
  • 3 years of experience in a technical leadership role.
  • 2 years of experience in a people management or team leadership role.
  • Experience with ML performance analysis, benchmarking, and computer architecture.

Responsibilities

  • Lead a team of software engineers focused on identifying and maintaining ML training and serving benchmarks that are representative to Google production and the broader ML industry.
  • Achieve performance for customer launches, and in case of third-party/open-source software (OSS) models, for engaged benchmark submissions (ML Commons, InferenceX, etc.).
  • Use benchmarks to identify performance opportunities and drive both near-term SOTA (e.g., custom kernels) and out-of the box performance (compiler/runtime optimizations, agentic tooling, auto-sharding) directly and in collaboration with partner teams.
  • Participate in algorithmic innovations exploiting new TPU hardware features and model-preserving optimizations (speculative decoding, sparsity, quantization, LoRA, etc.).
  • Participate in co-designing models that are TPU-friendly to showcase model quality at performance advanced to OSS models typically designed on GPUs.

Skills

Software development
ML design
Leadership
People management
ML performance analysis

Education

Bachelor’s degree or equivalent practical experience

Tools

CUDA
Triton
Pallas
MLIR/OpenXLA
PyTorch/JAX/vLLM

Job description

Google in Kirkland, WA is seeking an Engineering Manager for the Core ML/TPU Performance group. You will lead engineers across teams to optimize ML training and inference on TPU hardware, aligning software and hardware strategies with product goals and external collaborations.

The role emphasizes technical leadership, project budgeting, and multi-site coordination to deliver high-performance AI capabilities at scale. Applicants should have deep ML systems experience and people management skills.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Performance Engineering Manager
ML Performance Engineering Manager

Google • United States

On-site
USD 207,000 - 300,000
Health, dental, vision
401(k) with company match
Paid time off 20 days
+4
Engineering Manager: ML Performance & TPU Optimizations
Engineering Manager: ML Performance & TPU Optimizations

Socket.dev • Sunnyvale (CA)

On-site
USD 207,000 - 300,000
Engineering Manager, ML Performance
Engineering Manager, ML Performance

Socket.dev • Sunnyvale (CA)

On-site
USD 207,000 - 300,000
Engineering Manager, ML Performance
Engineering Manager, ML Performance

Google • Kirkland (WA)

On-site
USD 207,000 - 300,000
Health insurance
401(k) with company match
Paid Time Off: 20 days per year
+4
Staff ML Systems Co-Design Engineer
Staff ML Systems Co-Design Engineer

Socket.dev • Sunnyvale (CA)

On-site
USD 207,000 - 300,000
Staff ML Systems Co-Designer for TPU Performance
Staff ML Systems Co-Designer for TPU Performance

Google Inc. • Sunnyvale (CA)

On-site
USD 207,000 - 300,000
Health, dental, vision, life, and(dis)
401(k) with company match
Paid Time Off: 20 days/year
+4
Engineering Manager, ML Performance
Engineering Manager, ML Performance

Google • United States

On-site
USD 207,000 - 300,000
Health, dental, vision
401(k) with company match
Paid time off 20 days
+4
Engineering Manager, ML Performance
Engineering Manager, ML Performance

Google Inc. • Sunnyvale (CA)

On-site
USD 207,000 - 301,000
Health, dental, vision, life, and disability insurance
401(k) retirement benefits with company match
Paid Time Off: 20 days vacation per year
+1
Staff ML Systems Co-Design Engineer - Impact at Scale
Staff ML Systems Co-Design Engineer - Impact at Scale

Google • Sunnyvale (CA)

On-site
USD 207,000 - 300,000
Health, dental, vision, life, and disa
Senior ML Systems Architect – TPU & AI Infra
Senior ML Systems Architect – TPU & AI Infra

Socket.dev • Sunnyvale (CA)

On-site
USD 262,000 - 365,000
Equity
Benefits