Senior Manager, Compute Core Engineering

NVIDIA Gruppe

Santa Clara (CA)

On-site

USD 248,000 - 391,000

Full time

7 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA in Santa Clara, CA, seeks a Senior Manager to lead the Compute Core Engineering team and drive cutting-edge infrastructure across data centers, cloud, and compute platforms. You will build a culture of ownership, mentor engineers, and shape the roadmap for reliable, scalable services.

You will champion automation and AI-assisted operations, with hands-on scope in Go, Python, Terraform, and GitOps, and collaborate with leadership to deliver high-impact infrastructure products.

Qualifications

  • Bachelor's or Master's degree in a related technical field or equivalent experience.
  • 12+ years of experience with significant technical leadership; 5+ years in engineering management.
  • Proven track record crafting and managing critical infrastructure.
  • Strong distributed systems, networking, and Linux expertise.
  • Experience with container platforms, automation, and GitOps.

Responsibilities

  • Lead and expand NVIDIA's global Compute Core Engineering team with a culture of ownership and collaboration.
  • Define strategy and roadmap for infrastructure services across data centers, cloud, and compute platforms.
  • Drive automation and self-service platforms; modernize legacy infrastructure with minimal disruption.
  • Provide technical leadership for DNS, DHCP, NTP/PTP, LDAP, Linux infra, and related services.
  • Establish engineering standards for performance, security, capacity, and disaster recovery.
  • Champion AI-assisted operations and reduce manual work through automation.
  • Lead automation platforms with Go, Python, Terraform, and GitOps approaches.
  • Set service health metrics and lead incident reviews and reliability programs.
  • Forecast capacity and collaborate with teams to meet future demands.
  • Partner with NVIDIA leadership and customers to build compelling infrastructure products.

Skills

Leadership & mentoring
Distributed systems
Linux systems
Networking
Go / Python
GitOps
Automation
System performance
Communication to leadership

Education

Bachelor's or Master's in related field

Tools

Terraform
Go
Python
GitOps tooling

Job description

We are seeking a Senior Manager - Compute Core Engineering to lead our Compute Core Engineering team. This opportunity is outstanding as it allows you to contribute to groundbreaking advancements in AI and computing technology, advancing NVIDIA's legacy of innovation forward. You will have the chance to work with dedicated individuals and be part of a world-class team committed to flawless execution and ambitious goals.

What You Will Be Doing:
  • Lead and expand a group responsible for NVIDIA's global Compute Core Engineering Team. Recruit, mentor, and develop technical talent, encouraging a culture of ownership, collaboration, and innovation.
  • Define the strategy and roadmap for infrastructure services across NVIDIA's data centers, cloud environments, and compute platforms.
  • Drive the transformation of traditional services into automated, self-service platforms built for reliability and scalability. Modernize legacy infrastructure with minimal disruption.
  • Provide technical leadership for globally distributed services including DNS, DHCP, NTP/PTP, LDAP, and Linux infrastructure.
  • Establish engineering standards for performance, capacity, scalability, security, lifecycle management, and disaster recovery.
  • Build a culture centered around automation, operational excellence, and reducing manual work. Champion AI-assisted operations to automate diagnostics and routine infrastructure operations.
  • Partner with senior engineers to introduce technologies that improve infrastructure performance and resiliency. Lead the creation of automation platforms through Go, Python, Terraform, and GitOps approaches.
  • Establish metrics for service health, including availability, latency, capacity utilization, and infrastructure efficiency. Lead incident reviews and reliability improvement programs.
  • Lead forecasting activities and collaborate with various teams to ensure infrastructure can handle upcoming demands.
  • Collaborate closely with NVIDIA leadership and internal customers to build compelling infrastructure products.
What We Need To See:
  • Bachelor's or Master's degree in a related technical field or equivalent experience.
  • 12+ overall years of experience in related fields, including significant technical leadership experience. 5+ years of engineering management experience.
  • Demonstrated experience crafting and managing critically important infrastructure.
  • Solid grasp of distributed systems architecture. Solid knowledge of networking technologies and protocols.
  • Deep experience with Linux operating systems. Experience with container platforms and technologies.
  • Solid experience in infrastructure automation and GitOps approaches. Software engineering experience with Go, Python, or similar languages.
  • Experience using SLIs, SLOs, and performance metrics.
  • Experience in capacity forecasting and lifecycle oversight.
  • Demonstrated capability to guide intricate modernization projects.
  • Strong communication skills to translate technical topics for senior leadership. Demonstrated capability to draw in, champion, and keep high-achieving teams.
Ways To Stand Out From The Crowd:
  • Experience crafting and managing foundational infrastructure services. Experience supporting large AI, HPC, or accelerated computing environments.
  • Hands-on experience with advanced networking and infrastructure acceleration technologies. Deep comprehension of Linux kernel internals and improving system performance.
  • Experience modernizing legacy infrastructure into self-service platforms. Experience building engineering platforms around Terraform and GitOps or equivalent experience.
  • Experience implementing AI-assisted operations for infrastructure management. Proven track record of reducing operational toil through engineering and automation.
  • Experience leading globally distributed teams supporting critical infrastructure. Ability to remain technically engaged while providing organizational leadership and strategy.

NVIDIA is widely considered to be one of the technology world's most desirable employers. We have some of the most intelligent and hardworking people in the world working for us.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 248,000 USD - 391,000 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 3, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Manager, Compute Core Engineering
Senior Manager, Compute Core Engineering

NVIDIA • Santa Clara (CA)

On-site
USD 248,000 - 391,000
Equity
Benefits
Senior Manager, Compute Core Engineering
Senior Manager, Compute Core Engineering

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 248,000 - 391,000
Equity
Benefits
Senior Staff Site Reliability Engineer - Compute Core Engineering
Senior Staff Site Reliability Engineer - Compute Core Engineering

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

On-site
USD 200,000 - 322,000
Principal Software Engineer - Compute Infrastructure
Principal Software Engineer - Compute Infrastructure

Nvidia Corporation in • Santa Clara (CA)

Hybrid
USD 248,000 - 391,000
Equity
Benefits
Senior Staff Site Reliability Engineer - Compute Core Engineering
Senior Staff Site Reliability Engineer - Compute Core Engineering

Socket.dev • California (MO)

Hybrid
USD 200,000 - 322,000
Equity
Benefits
Senior Infrastructure Solutions Architect
Senior Infrastructure Solutions Architect

NVIDIA • California (MO)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Manager, Business Operations
Senior Manager, Business Operations

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 240,000 - 380,000
Senior Staff Platform Engineer
Senior Staff Platform Engineer

Nvidia Corporation • Santa Clara (CA)

On-site
USD 200,000 - 322,000
Equity
Comprehensive benefits
Senior Infrastructure Solutions Architect
Senior Infrastructure Solutions Architect

Socket.dev • Austin (TX)

On-site
USD 152,000 - 288,000
Senior Software Engineer, Infrastructure Automation and Distributed Systems
Senior Software Engineer, Infrastructure Automation and Distributed Systems

NVIDIA Gruppe • United States

On-site
USD 230,000 - 430,000
Equity
Benefits