Director, Technical Program Management, ML Infrastructure Deployment

Google Inc.

Mountain View (CA)

On-site

USD 281,000 - 392,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Benefits

Job summary

Google Mountain View, CA, USA is seeking a Director of Technical Program Management for ML Infrastructure Deployment. You will lead planning and execution of ML infrastructure deployments across product areas and external customers, ensuring scalable, efficient data-center operations.

This role emphasizes mentoring a team of execution leads for TPU/GPU platforms, defining KPIs, and collaborating with data center leads to meet demand and accelerate delivery.

Qualifications

  • Bachelor's degree in Engineering or equivalent practical experience
  • 15 years of experience in ML infrastructure planning or deployment - in NPI or production

Responsibilities

  • Develop executable deployment plans for Google's ML infrastructure, ensuring efficient use of power and cooling at data center level.
  • Drive execution of ML infrastructure deployments to meet Google's demand and ensure on-time delivery of capacity.
  • Lead, mentor, and develop a high-performing team of execution leads for TPU/GPU platforms; manage staff, budget, and resources.
  • Lead and influence cross-functional ML workstreams in a matrix environment to deliver acceleration outcomes.
  • Establish KPIs and metrics to monitor and improve the effectiveness of planning and execution.

Job description

Director, Technical Program Management, ML Infrastructure Deployment

Google Mountain View, CA, USA

  • Bachelor's degree in Engineering or equivalent practical experience
  • 15 years of experience in ML infrastructure planning or deployment - in NPI or production
Preferred qualifications:
  • Master's degree in Engineering
  • Solid understanding of infrastructure deployment processes, dependencies, constraints and acceleration approaches
About the job

Google’s AI & Infrastructure (AI2) team is responsible for the global infrastructure that powers Google. We do everything from designing and building gigawatts of data centers and petabits of networks around the world, to designing and manufacturing the servers, AI accelerators, and optical switches which comprise our fleet, to developing the infrastructure software which runs on that fleet and underpins all Google’s multi-billion-user services and Cloud platforms.

Our organization delivers sustainable and resource-efficient infrastructure that supports Google and our customers. We provide a foundation that operates safely, securely, and sustainably, and are chartered with ensuring Google’s data center landscape can support rapidly changing demand while delivering on Google’s commitment to sustainability.

We seek a leader that will be responsible for driving execution of current ML infrastructure deployments and working closely with all partner teams involved.

In this key leadership role, you will be a key lead responsible for managing the execution and orchestration of ML infrastructure deployment for Google Product Areas and external customers. In addition, you’ll be a key stakeholder in ensuring that our ML infrastructure planning is robust and comprehensive.

Key responsibilities
  • Work with planning leads to develop executable and deployment plans for Google’s ML infrastructure. Ensure ML infrastructure plans make efficient use of power and cooling at a data center level.
  • Drive execution of ML infrastructure deployments, ensuring timely delivery to meet Google's demand. Oversee execution escalations to ensure timely closure of issues and on-time delivery of capacity.
  • Lead, mentor, and develop a high-performing team of execution leads for TPU/GPU platforms. Manage and allocate staff, budget, and resources effectively to achieve key objectives.
  • Lead and influence cross-functional ML workstreams in a matrix environment to deliver outstanding acceleration outcomes.
  • Work closely with leads in functional areas and data centers to drive scalable execution and acceleration at a multi-GW scale. Establish key performance indicators (KPIs) and metrics to monitor and improve the effectiveness of planning and execution.

US: $281000 - $392000 (USD) + 30% bonus target + equity + benefits

Google is proud to be an equal opportunity and affirmative action employer. We are committed to building a workforce that is representative of the users we serve, creating a culture of belonging, and providing an equal employment opportunity regardless of race, creed, color, religion, gender, sexual orientation, gender identity/expression, national origin, disability, age, genetic information, veteran status, marital status, pregnancy or related condition (including breastfeeding), expecting or parents-to-be, criminal histories consistent with legal requirements, or any other basis protected by law. See also Google's EEO Policy , Know your rights: workplace discrimination is illegal , Belonging at Google , and How we hire .

Google is a global company and, in order to facilitate efficient collaboration and communication globally, English proficiency is a requirement for all roles unless stated otherwise in the job posting.

Equity is granted exclusively and discretionarily by Alphabet Inc. on the basis of an agreement concluded between you and Alphabet Inc. Alphabet Inc. is your sole contractual partner with respect to equity grants. GSU grants are not guaranteed, are discretionary, are subject to approval by the Alphabet Inc. board of directors or its delegate, the terms of the relevant Alphabet Inc. stock plan, and your grant agreement. They have no impact on statutory payments. Current or past grants do not confer an acquired right.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer, ML Infrastructure, Control Plane
Senior Software Engineer, ML Infrastructure, Control Plane

Google • Sunnyvale (CA)

On-site
USD 174,000 - 252,000
Technical Program Manager II, Capacity Constraints Management, Cloud Infrastructure
Technical Program Manager II, Capacity Constraints Management, Cloud Infrastructure

Google Inc. • Atlanta (GA)

On-site
USD 138,000 - 197,000
Technical Program Manager, ML Efficiency
Technical Program Manager, ML Efficiency

Google • Sunnyvale (CA)

On-site
USD 192,000 - 278,000
Senior Staff Software Engineer, AI/ML, Google Cloud
Senior Staff Software Engineer, AI/ML, Google Cloud

Google Inc. • Seattle (WA)

On-site
USD 262,000 - 365,000
25% bonus target
Equity opportunities
Comprehensive benefits package
Engineering Manager, ML Performance
Engineering Manager, ML Performance

Google • United States

On-site
USD 207,000 - 300,000
Health, dental, vision
401(k) with company match
Paid time off 20 days
+4
Technical Program Manager, ML Efficiency
Technical Program Manager, ML Efficiency

Google Inc. • Sunnyvale (CA)

On-site
USD 192,000 - 278,000
Tech Lead Manager, Staff Software Engineering, XProf
Tech Lead Manager, Staff Software Engineering, XProf

Google • Sunnyvale (CA)

On-site
USD 207,000 - 301,000
Senior Technical Program Manager II, AI/ML Systems
Senior Technical Program Manager II, AI/ML Systems

Google • Sunnyvale (CA)

On-site
USD 200,000 - 333,000
Staff Software Engineer, Network Health
Staff Software Engineer, Network Health

Google Inc. • Sunnyvale (CA)

On-site
USD 207,000 - 301,000
20% Bonus Target
Equity Options
Comprehensive Benefits
Technical Program Manager II, Capacity Constraints Management, Cloud Infrastructure
Technical Program Manager II, Capacity Constraints Management, Cloud Infrastructure

Google • Atlanta (GA)

On-site
USD 138,000 - 197,000