Senior Datacenter Technical Program Manager, At-Scale AI Clusters

NVIDIA

United States

On-site

USD 168,000 - 322,000

Full time

29 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA is seeking a Technical Program Manager to lead datacenter integration for next-generation AI supercomputing systems. You will oversee the lifecycle from design and requirements to production deployment and support, coordinating across engineering teams and external partners.

The role emphasizes HP computing expertise, strong teamwork, and effective communication with leadership to ensure successful deployments for large customers. Equity and benefits are included.

Qualifications

  • BS in Applied Science or Engineering (or equivalent experience).
  • 8+ years of overall experience.
  • Experience with high-performance computing systems and GPU clusters deployed in on-premises datacenters.
  • Strong teamwork and interpersonal skills to facilitate coordinated workflows.

Responsibilities

  • Collaborate with engineers and architects to build and deploy large scale GPU computing systems.
  • Lead the integration of new AI clusters with datacenter facilities with demanding power, cooling, and instrumentation requirements.
  • Coordinate design and fit-out of new datacenter builds with internal teams and external contractors.
  • Own and produce detailed documentation for the end-to-end datacenter fit-out and integration process.
  • Communicate with engineering leadership to prioritize issues essential to major customers.

Skills

HPC systems
GPU clusters
Team collaboration

Education

BS in Applied Science or Engineering

Job description

NVIDIA is looking for a highly-motivated Technical Program Manager (TPM) to join our Applied Systems Engineering Team to drive datacenter integration for the next generation of NVIDIA AI supercomputing systems. This TPM will play a crucial role throughout the lifecycle of the latest AI systems at scale, from datacenter design and requirements definition, through systems integration of AI clusters into the datacenter environment, and support for these systems as they enter production.

What You’ll Be Doing
  • Collaborate with outstanding engineers and architects to build and deploy large scale GPU computing systems based on NVIDIA's reference supercomputing architectures
  • Lead the integration of new AI clusters with datacenter facilities with demanding requirements on power, cooling, and instrumentation
  • Coordinate design and fit-out of new datacenter builds, working with both internal engineering teams and external contractors
  • Own and produce detailed documentation for the end-to-end process for datacenter fit-out and integration
  • Communicate internally with engineering leadership to prioritize and address key issues essential to the success of our largest customers
What We Need To See
  • BS in Applied Science or Engineering (or equivalent experience)
  • 8+ years of overall experience
  • Experience with high-performance computing systems and GPU clusters deployed in on-premises datacenters
  • A passion for understanding challenging technical problems and driving the process of finding a solution
  • Strong teamwork and interpersonal skills, to facilitate building a collaborative workflow for coordination between many teams
Ways To Stand Out From The Crowd
  • Understanding of datacenter design, including familiarity with power and cooling technologies
  • Expertise in system monitoring and instrumentation of large clusters, using technologies such as Prometheus, Grafana, Splunk, Modbus, and BACNet
  • Experience working with the engineering or academic research community supporting high-performance computing or deep learning

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 168,000 USD - 258,750 USD for Level 4, and 200,000 USD - 322,000 USD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 4, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Datacenter Technical Program Manager, At-Scale AI Clusters
Senior Datacenter Technical Program Manager, At-Scale AI Clusters

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 168,000 - 322,000
Equity
Benefits
Senior Solutions Architect, AI Compute – NPN
Senior Solutions Architect, AI Compute – NPN

NVIDIA Gruppe • California (MO)

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior Technical Program Manager, Data Center - Engineering Operations
Senior Technical Program Manager, Data Center - Engineering Operations

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 168,000 - 258,750
Equity
Comprehensive benefits
Senior Solutions Architect, AI Compute – NPN
Senior Solutions Architect, AI Compute – NPN

NVIDIA Corporation • Northern (KY)

On-site
USD 184,000 - 357,000
Equity
Benefits package
Senior Infrastructure Solutions Architect
Senior Infrastructure Solutions Architect

Socket.dev • Austin (TX)

On-site
USD 152,000 - 288,000
Senior Infrastructure Solutions Architect
Senior Infrastructure Solutions Architect

NVIDIA • California (MO)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior HPC AI Cluster Engineer
Senior HPC AI Cluster Engineer

NVIDIA • California (MO)

On-site
USD 176,000 - 334,000
Senior HPC AI Cluster Engineer
Senior HPC AI Cluster Engineer

NVIDIA • Santa Clara (CA)

On-site
USD 176,000 - 334,000
Equity
Benefits
Senior Systems Software Engineer - GPU Performance at Scale
Senior Systems Software Engineer - GPU Performance at Scale

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Benefits
Senior Infrastructure Solutions Architect
Senior Infrastructure Solutions Architect

NVIDIA • Austin (TX)

On-site
USD 152,000 - 288,000
Equity
Benefits