Engineering Manager, DSC AI Inference Platform

Google

Sunnyvale (CA)

On-site

USD 207,000 - 301,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Google Sunnyvale is seeking an Engineering Manager for the DSC AI Inference Platform to lead a team of systems and ML engineers. You will shape technical goals, manage delivery, and mentor engineers across multiple sites and disciplines.

You will drive architecture for disaggregated serving, optimize LLM performance on GPU accelerators, and partner with Research, SRE, and Core GPU teams to deploy AI models at scale.

Qualifications

  • Bachelor's degree or equivalent practical experience.
  • 8 years of experience programming in C++ or Python.
  • 5 years of experience optimizing, profiling, and scaling production‑grade systems on GPU accelerators or specialized AI hardware.
  • 5 years of experience directly managing and leading engineering teams focused on machine learning infrastructure, AI platforms, or high‑performance distributed computing systems.
  • 5 years of experience in a people management or team leadership role.
  • 3 years of experience managing engineering organizations across multi‑team infrastructure dependencies.

Responsibilities

  • Lead, mentor, and grow a high‑performing team of systems and ML engineers across multiple sites.
  • Define the technical goal and strategy for enhancing the LLM serving stack, focusing on performance, scalability, and resource efficiency.
  • Drive the design and implementation of advanced serving architectures, including disaggregated serving, to optimize resource utilization and latency.
  • Oversee the building and maintenance of critical infrastructure and tooling for in‑depth performance analysis, profiling, and benchmarking of LLM models on GPU accelerators.
  • Partner with Research, SRE, Product, and Core GPU library teams to optimize and deploy LLMs in production globally.

Skills

C++
Python
GPU optimization
ML infrastructure
People management
Multi-team leadership
Engineering leadership

Education

Bachelor's degree or equivalent practical experience
Master’s or PhD preferred

Tools

Nsight
xprof
JAX
PyTorch
TensorFlow
CUDA

Job description

Engineering Manager, DSC AI Inference Platform

Company: Google – Sunnyvale, CA, USA

Advanced

Experience owning outcomes and decision making, solving ambiguous problems and influencing stakeholders; deep expertise in domain.

Minimum qualifications:
  • Bachelor's degree or equivalent practical experience.
  • 8 years of experience programming in C++ or Python.
  • 5 years of experience optimizing, profiling, and scaling production‑grade systems on GPU accelerators or specialized AI hardware.
  • 5 years of experience directly managing and leading engineering teams focused on machine learning infrastructure, AI platforms, or high‑performance distributed computing systems.
  • 5 years of experience in a people management or team leadership role.
  • 3 years of experience managing engineering organizations across multi‑team infrastructure dependencies.
Preferred qualifications:
  • Master's degree or PhD degree in Computer Science or a related technical field.
  • 5 years of experience working in a complex, matrixed organization.
  • 4 years of experience implementing advanced LLM serving architectures and optimization techniques, such as disaggregated serving, continuous batching, or specialized compiler technologies (e.g., XLA).
  • 3 years of experience utilizing deep‑dive ML profiling tools (e.g., Nsight, xprof) to troubleshoot and resolve low‑level bottlenecks within major frameworks like JAX, PyTorch, or TensorFlow.
About the job

Like Google's own ambitions, the work of a Software Engineer goes beyond just Search. Software Engineering Managers have not only the technical expertise to take on and provide technical leadership to major projects, but also manage a team of Engineers. You not only optimize your own code but make sure Engineers are able to optimize theirs. As a Software Engineering Manager you manage your project goals, contribute to product strategy and help develop your team. Teams work all across the company, in areas such as information retrieval, artificial intelligence, natural language processing, distributed computing, large‑scale system design, networking, security, data compression, user interface design; the list goes on and is growing every day. Operating with scale and speed, our exceptional software engineers are just getting started — and as a manager, you guide the way.

With technical and leadership expertise, you manage engineers across multiple teams and locations, a large product budget and oversee the deployment of large‑scale projects across multiple sites internationally.

The Distributed Cloud (DSC) AI Inference Platform team operates at the critical intersection of Large Language Models (LLMs) and high‑performance computing. Our mission is to engineer the future of AI serving infrastructure, driving foundational improvements in efficiency, latency, and throughput. We develop innovative solutions, including disaggregated serving architectures, and build the essential tools to analyze and optimize LLM performance on cutting‑edge GPU platforms. Our work directly enables Google to deploy and scale state‑of‑the‑art AI models (like Gemini) effectively and efficiently across Google's global infrastructure, products, and Cloud.

The Google Cloud AI Research team addresses AI challenges motivated by Google Cloud’s mission of bringing AI to tech, healthcare, finance, retail and many other industries. We work on a range of unique problems focused on research topics that maximize scientific and real‑world impact, aiming to push the state‑of‑the‑art in AI and share findings with the broader research community. We also collaborate with product teams to bring innovations to real‑world impact that benefits our customers. Individual pay is determined by factors including job‑related skills, experience, and relevant education or training.

US: $207,000 - $301,000 (USD) + 20% bonus target + equity + benefits.

Learn more about benefits at Google.

Responsibilities
  • People Management and Talent Development: Lead, mentor, and grow a high‑performing team of systems and ML engineers. Drive a culture of excellence, psychological safety, and continuous learning. Guide career paths, define OKRs, and conduct performance evaluations.
  • Strategic and Technical Roadmap: Define the technical goal and strategy for enhancing the LLM serving stack, focusing on performance, scalability, and resource efficiency.
  • Architectural Leadership: Drive the design and implementation of advanced serving architectures, including disaggregated serving, to optimize resource utilization and latency.
  • Infrastructure Oversight: Oversee the building and maintenance of critical infrastructure and tooling for in‑depth performance analysis, profiling, and benchmarking of LLM models on GPU accelerators.
  • Cross-Functional Collaboration: Partner closely with Research, SRE, Product, and Core GPU library teams to optimize and deploy LLMs in production globally. Align team efforts with broader organizational AI priorities.

Information collected and processed as part of your Google Careers profile, and any job applications you choose to submit is subject to Google's Applicant and Candidate Privacy Policy.

Google is proud to be an equal opportunity and affirmative action employer. We are committed to building a workforce that is representative of the users we serve, creating a culture of belonging, and providing an equal employment opportunity regardless of race, creed, color, religion, gender, sexual orientation, gender identity/expression, national origin, disability, age, genetic information, veteran status, marital status, pregnancy or related condition (including breastfeeding), expecting or parents‑to‑be, criminal histories consistent with legal requirements, or any other basis protected by law.

If you have a need that requires accommodation, please let us know by completing our Accommodations for Applicants form.

Google is a global company and, in order to facilitate efficient collaboration and communication globally, English proficiency is a requirement for all roles unless stated otherwise in the job posting.

To all recruitment agencies: Google does not accept agency resumes. Please do not forward resumes to our jobs alias, Google employees, or any other organization location. Google is not responsible for any fees related to unsolicited resumes.

Equity is granted exclusively and discretionarily by Alphabet Inc. on the basis of an agreement concluded between you and Alphabet Inc. Alphabet Inc. is your sole contractual partner with respect to equity grants. GSU grants are not guaranteed, are discretionary, are subject to approval by the Alphabet Inc. board of directors or its delegate, the terms of the relevant Alphabet Inc. stock plan, and your grant agreement. They have no impact on statutory payments. Current or past grants do not confer an acquired right.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Staff Engineer, GDC AI Inference Platform
Senior Staff Engineer, GDC AI Inference Platform

Google • United States

On-site
USD 262,000 - 365,000
Health insurance
401(k) match
Paid time off
+4
Tech Lead Manager, Staff Software Engineering, XProf
Tech Lead Manager, Staff Software Engineering, XProf

Google • Sunnyvale (CA)

On-site
USD 207,000 - 301,000
Senior Software Engineering Manager, AI/ML, Google Cloud AI
Senior Software Engineering Manager, AI/ML, Google Cloud AI

Google Inc. • Sunnyvale (CA)

On-site
USD 262,000 - 365,000
Senior Software Engineer, Google Distributed Cloud AI
Senior Software Engineer, Google Distributed Cloud AI

Google • Sunnyvale (CA)

On-site
USD 174,000 - 252,000
Software Engineering Manager II, AI/ML GenAI, Google Cloud Applications AI
Software Engineering Manager II, AI/ML GenAI, Google Cloud Applications AI

Google • Sunnyvale (CA)

On-site
USD 207,000 - 300,000
Software Engineering Manager II, AI/ML GenAI, Google Cloud Compute
Software Engineering Manager II, AI/ML GenAI, Google Cloud Compute

Google • Kirkland (WA)

On-site
USD 207,000 - 301,000
Health, dental, vision, life, disability insurance
401(k) with company match
20 days vacation per year
+3
Engineering Director, Spatial Flex
Engineering Director, Spatial Flex

Google • Sunnyvale (CA)

On-site
USD 307,000 - 427,000
Health insurance
401(k) with company match
Paid Time Off: 20 days/year
+4
Software Engineering Manager II, AI/ML GenAI, Google Cloud AI
Software Engineering Manager II, AI/ML GenAI, Google Cloud AI

Google • Kirkland (WA)

On-site
USD 207,000 - 301,000
Health insurance
Dental insurance
Vision insurance
+8
Senior Staff Engineer, GDC AI Inference Platform
Senior Staff Engineer, GDC AI Inference Platform

Google Inc. • Sunnyvale (CA)

On-site
USD 262,000 - 365,000
Health insurance
Dental insurance
Vision insurance
+7
Engineering Manager, ML Performance
Engineering Manager, ML Performance

Google • United States

On-site
USD 207,000 - 300,000
Health, dental, vision
401(k) with company match
Paid time off 20 days
+4