ML Inference & Compiler Engineer (Accelerators)

Google DeepMind

Mountain View (CA)

On-site

USD 174,000 - 252,000

Full time

12 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Google DeepMind is building a next-generation compute platform for AI inference workloads, with memory tech, HW acceleration, and SW/HW codesign. We seek engineers who take initiative to advance ML compiler stack and inference infrastructure, joining a team that directly impacts AI development for Google and industry.

This is a coding role for engineers who enjoy hands-on work, collaboration, and delivering high-performance software.

Qualifications

  • Bachelor's degree in Computer Science, a related technical field, or equivalent practical experience.
  • Experience in software development using Python and C++.
  • Experience with ML hardware accelerators, ML compiler stacks (e.g., JAX, PyTorch, XLA), and kernel development.

Responsibilities

  • Direct full-stack Software role focusing on ML compiler and inference infrastructure.
  • Design and develop ML compiler stack to bridge AI workloads and low-level Hardware operators.
  • Design and develop SW infrastructure to enable high performance inference.
  • Drive SW/HW codesign, kernel optimization, performance tuning, and debugging.

Skills

Python
C++
ML accelerators
ML compilers
Kernel development

Education

Bachelor's degree in Computer Science, related field, or equivalent practical experience
PhD degree in Computer Engineering/CS (preferred)

Tools

JAX
PyTorch
XLA

Job description

Google DeepMind is building a next-generation compute platform for AI inference workloads, with memory tech, HW acceleration, and SW/HW codesign. We seek engineers who take initiative to advance ML compiler stack and inference infrastructure, joining a team that directly impacts AI development for Google and industry.

This is a coding role for engineers who enjoy hands-on work, collaboration, and delivering high-performance software.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Accelerators Software Engineer – Inference & Compiler
ML Accelerators Software Engineer – Inference & Compiler

Google Inc. • Mountain View (CA), San Francisco (CA)

On-site
USD 174,000 - 252,000
ML Inference & HW Codesign Engineer
ML Inference & HW Codesign Engineer

Socket.dev • Mountain View (CA)

On-site
USD 174,000 - 252,000
ML Inference & Accelerator Software Engineer
ML Inference & Accelerator Software Engineer

Google DeepMind • San Francisco (CA)

On-site
USD 174,000 - 252,000
Equity
Benefits
Software Engineer, ML Accelerators, DeepMind
Software Engineer, ML Accelerators, DeepMind

Socket.dev • Mountain View (CA)

On-site
USD 174,000 - 252,000
AI Accelerator Engineer (HW/SW Co-Design) – Equity
AI Accelerator Engineer (HW/SW Co-Design) – Equity

Socket.dev • Mountain View (CA)

On-site
USD 174,000 - 252,000
Bonus target 15%
AI Compute Engineer — Accelerators Co-Design
AI Compute Engineer — Accelerators Co-Design

Google • United States

On-site
USD 174,000 - 252,000
AI Inference Compute Engineer — Hardware/Software Co-Design
AI Inference Compute Engineer — Hardware/Software Co-Design

Google DeepMind • San Francisco (CA)

On-site
USD 174,000 - 252,000
Accelerator Software Engineer: HW/SW Co-Design
Accelerator Software Engineer: HW/SW Co-Design

Google DeepMind • Mountain View (CA)

On-site
USD 174,000 - 252,000
Bonus target 15%
Equity
Benefits
Software Engineer, ML Accelerators, DeepMind
Software Engineer, ML Accelerators, DeepMind

Google DeepMind • San Francisco (CA)

On-site
USD 174,000 - 252,000
Equity
Benefits
Software Engineer, ML Accelerators, DeepMind
Software Engineer, ML Accelerators, DeepMind

Google DeepMind • Mountain View (CA)

On-site
USD 174,000 - 252,000