Technical Program Manager, AI Factory Infrastructure

NVIDIA AI

Santa Clara (CA)

On-site

USD 240,000 - 380,000

Full time

2 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA AI is seeking a Program Manager for its AI Factory Infrastructure team in Santa Clara, CA. You will lead multi-disciplinary teams to design, build, and deploy 100MW+ AI factory data centers, coordinating with electrical, mechanical, and network groups from planning through handoff to operations.

You will drive long-term programs, manage resources and schedules, and ensure alignment with engineering roadmaps and government initiatives. A senior PM with data center experience is preferred.

Qualifications

  • 12+ years of experience providing program and project management leadership for data center projects covering construction of mechanical, electrical, and plumbing with large-scale server, storage, and network deployments.
  • BS or MS degree in Engineering (or equivalent experience).
  • Experience managing end-to-end data center deployments for high-density AI infrastructure, including commissioning, readiness reviews, turn-up, and operational handoff.
  • In-depth knowledge of infrastructure data center facilities infrastructure (electrical and mechanical) technologies.

Responsibilities

  • Collaborate with product owners and technical leads to identify and collect requirements for next-generation AI Factories.
  • Build, supervise, and complete long-term programs including schedules, resourcing, and checkpoints.
  • Work with data center and hardware teams to find creative solutions to hard problems, and co-develop solutions and mitigation strategies.
  • Lead planning with key internal partners on capacity demands with engineering roadmaps and data center expansions.
  • Own end-to-end delivery of 100MW+ AI factory data center deployments, from construction readiness through commissioning, turn-up, and handoff to operations.
  • Coordinate general contractors, colocation providers, utilities, and OEMs to align electrical/mechanical scope, long-lead equipment, and site logistics for large-scale AI cluster deployments.
  • Drive integrated readiness reviews and acceptance criteria across power, liquid cooling, networking, and platform/hardware teams to ensure performance and reliability targets are met for AI factory applications.
  • Develop program plans for government grants and initiatives.
  • Translate program requirements into Basis Of Design documents
  • Bring together team members and foster a collaborative approach to delivery while holding team members accountable to action items and timelines.

Skills

Program management
Data center planning
Cross-functional leadership
Power delivery
Project scheduling

Education

BS or MS in Engineering

Job description

Job Requisition ID JR2024574

Job Category Program Manager

Time Type Full time

NVIDIA's AI Factory Infrastructure team develops global reference designs, de-risks technologies needed for the next generation of compute and network products and builds AI infrastructure at scale to validate solutions at scale. As a TPM on the team, you’ll be responsible for leading teams that are highly multi-functional, including hardware, software and facility infrastructure teams, to deliver solutions and infrastructure. This role offers a truly unique blend of developing future solutions and products with building infrastructure at scale.

What You Will Be Doing
  • Collaborate with product owners and technical leads to identify and collect requirements for next-generation AI Factories.
  • Build, supervise, and complete long-term programs including schedules, resourcing, and checkpoints.
  • Work with data center and hardware teams to find creative solutions to hard problems, and co-develop solutions and mitigation strategies.
  • Lead planning with key internal partners on capacity demands with engineering roadmaps and data center expansions.
  • Own end-to-end delivery of 100MW+ AI factory data center deployments, from construction readiness through commissioning, turn-up, and handoff to operations.
  • Coordinate general contractors, colocation providers, utilities, and OEMs to align electrical/mechanical scope, long-lead equipment, and site logistics for large-scale AI cluster deployments.
  • Drive integrated readiness reviews and acceptance criteria across power, liquid cooling, networking, and platform/hardware teams to ensure performance and reliability targets are met for AI factory applications.
  • Develop program plans for government grants and initiatives.
  • Translate program requirements into Basis Of Design documents
  • Bring together team members and foster a collaborative approach to delivery while holding team members accountable to action items and timelines.
What We Need To See
  • Outstanding long-term planning and execution skills to carry our data center lifecycle planning, including large-scale AI factory buildouts and expansions.
  • Experience managing end-to-end data center deployments for high-density AI infrastructure, including commissioning, readiness reviews, turn-up, and operational handoff.
  • Demonstrated ability to coordinate across colocation providers, general contractors, utilities, and OEMs to deliver complex electrical and mechanical scope (e.g., at 100MW+ campus scale).
  • Strong technical and program leadership across power delivery, liquid cooling, networking, and compute/platform teams to define acceptance criteria and ensure performance and reliability targets are met.
  • 12+ years of experience providing program and project management leadership for data center projects covering construction of mechanical, electrical, and plumbing with large-scale server, storage, and network deployments.
  • BS or MS degree in Engineering (or equivalent experience).
Ways To Stand Out From The Crowd
  • In-depth knowledge of infrastructure (hardware and software) data center facilities infrastructure (electrical and mechanical) technologies.
  • Familiarity with NVIDIA’s AI compute technology stack and ability to translate platform requirements into data center infrastructure designs (power delivery, liquid cooling, space, and network topology) at scale.
  • Experience with colocation data center environments

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 200,000 USD - 322,000 USD for Level 5, and 240,000 USD - 379,500 USD for Level 6.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 24, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Solutions Architect, Data Center Buildouts – NPN
Senior Solutions Architect, Data Center Buildouts – NPN

NVIDIA Gruppe • California (MO)

On-site
USD 210,000 - 357,000
Equity
Benefits
Senior Datacenter Technical Program Manager, At-Scale AI Clusters
Senior Datacenter Technical Program Manager, At-Scale AI Clusters

NVIDIA • California (MO)

On-site
USD 168,000 - 322,000
Equity and Benefits
Senior Solutions Architect, Data Center Buildouts – NPN
Senior Solutions Architect, Data Center Buildouts – NPN

NVIDIA Corporation • Northern (KY)

Hybrid
USD 184,000 - 357,000
Equity
Senior Datacenter Technical Program Manager, At-Scale AI Clusters
Senior Datacenter Technical Program Manager, At-Scale AI Clusters

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 168,000 - 322,000
Equity
Benefits
Senior Datacenter Technical Program Manager, At-Scale AI Clusters
Senior Datacenter Technical Program Manager, At-Scale AI Clusters

NVIDIA • United States

On-site
USD 168,000 - 322,000
Equity
Benefits
Director, Technical Program Management, Enterprise & AI
Director, Technical Program Management, Enterprise & AI

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 272,000 - 426,000
Solutions Architect, Data Center Buildouts
Solutions Architect, Data Center Buildouts

NVIDIA Gruppe • Town of Texas (WI)

On-site
USD 224,000 - 431,000
Equity
Benefits
Technical Product Manager - AI Infra Resilience
Technical Product Manager - AI Infra Resilience

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 240,000 - 380,000
Equity
Benefits
Senior Solutions Architect, Data Center Buildouts – NPN
Senior Solutions Architect, Data Center Buildouts – NPN

NVIDIA • California (MO)

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior Solutions Architect, Data Center Buildouts – NPN
Senior Solutions Architect, Data Center Buildouts – NPN

NVIDIA • United States

On-site
USD 184,000 - 357,000
Equity
Benefits