Senior Manager, Compute Core Engineering

NVIDIA

Santa Clara (CA)

On-site

USD 248,000 - 391,000

Full time

7 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA seeks a Senior Manager to lead Compute Core Engineering, guiding a global team to build scalable, automated infrastructure across data centers and clouds.

You will define roadmaps, modernize legacy systems, and drive AI-enabled operations while partnering with senior engineering to improve performance and resiliency. Strong leadership, Linux expertise, and go/python skills are essential.

Qualifications

  • Bachelor's or Master's degree in a technical field or equivalent experience.
  • 12+ years in related fields with significant technical leadership, 5+ years in engineering management.
  • Experience crafting and managing foundational infrastructure services.
  • Strong networking knowledge and understanding of protocols.
  • Deep experience with Linux and container platforms.
  • Automation and GitOps experience; software engineering in Go or Python.
  • Experience with SLIs, SLOs, and performance metrics.
  • Forecasting and lifecycle oversight experience.
  • Ability to lead modernization projects and communicate with leadership.

Responsibilities

  • Lead and expand the global Compute Core Engineering Team; recruit, mentor, and develop talent.
  • Define strategy and roadmap for infrastructure services across data centers, cloud, and compute platforms.
  • Transform services into automated, self-service platforms and modernize legacy infrastructure.
  • Provide technical leadership for DNS, DHCP, NTP/PTP, LDAP, and Linux infra.
  • Set engineering standards for performance, capacity, scalability, security, and DR.
  • Drive automation and AI-assisted operations to reduce manual toil.
  • Collaborate with senior engineers to introduce technologies for performance and resiliency; build automation platforms with Go, Python, Terraform, GitOps.
  • Establish service health metrics and lead incident reviews and reliability programs.
  • Forecast capacity and collaborate with teams to meet future demands.
  • Work with leadership and customers to build compelling infrastructure products.

Skills

Distributed systems
Networking
Linux
Container platforms
Go
Python
GitOps
Incident management

Education

Bachelor's or Master's degree in a related field

Tools

Go
Python

Job description

We are seeking a Senior Manager - Compute Core Engineering to lead our Compute Core Engineering team. This opportunity is outstanding as it allows you to contribute to groundbreaking advancements in AI and computing technology, advancing NVIDIA's legacy of innovation forward. You will have the chance to work with dedicated individuals and be part of a world-class team committed to flawless execution and ambitious goals.

What You Will Be Doing:
  • Lead and expand a group responsible for NVIDIA's global Compute Core Engineering Team. Recruit, mentor, and develop technical talent, encouraging a culture of ownership, collaboration, and innovation.

  • Define the strategy and roadmap for infrastructure services across NVIDIA's data centers, cloud environments, and compute platforms.

  • Drive the transformation of traditional services into automated, self-service platforms built for reliability and scalability. Modernize legacy infrastructure with minimal disruption.

  • Provide technical leadership for globally distributed services including DNS, DHCP, NTP/PTP, LDAP, and Linux infrastructure.

  • Establish engineering standards for performance, capacity, scalability, security, lifecycle management, and disaster recovery.

  • Build a culture centered around automation, operational excellence, and reducing manual work. Champion AI-assisted operations to automate diagnostics and routine infrastructure operations.

  • Partner with senior engineers to introduce technologies that improve infrastructure performance and resiliency. Lead the creation of automation platforms through Go, Python, Terraform, and GitOps approaches.

  • Establish metrics for service health, including availability, latency, capacity utilization, and infrastructure efficiency. Lead incident reviews and reliability improvement programs.

  • Lead forecasting activities and collaborate with various teams to ensure infrastructure can handle upcoming demands.

  • Collaborate closely with NVIDIA leadership and internal customers to build compelling infrastructure products.

What We Need To See:
  • Bachelor's or Master's degree in a related technical field or equivalent experience.

  • 12+ overall years of experience in related fields, including significant technical leadership experience. 5+ years of engineering management experience.

  • Demonstrated experience crafting and managing critically important infrastructure.

  • Solid grasp of distributed systems architecture. Solid knowledge of networking technologies and protocols.

  • Deep experience with Linux operating systems. Experience with container platforms and technologies.

  • Solid experience in infrastructure automation and GitOps approaches. Software engineering experience with Go, Python, or similar languages.

  • Experience using SLIs, SLOs, and performance metrics.

  • Experience in capacity forecasting and lifecycle oversight.

  • Demonstrated capability to guide intricate modernization projects.

  • Strong communication skills to translate technical topics for senior leadership. Demonstrated capability to draw in, champion, and keep high-achieving teams.

Ways To Stand Out From The Crowd:
  • Experience crafting and managing foundational infrastructure services. Experience supporting large AI, HPC, or accelerated computing environments.

  • Hands-on experience with advanced networking and infrastructure acceleration technologies. Deep comprehension of Linux kernel internals and improving system performance.

  • Experience modernizing legacy infrastructure into self-service platforms. Experience building engineering platforms around Terraform and GitOps or equivalent experience.

  • Experience implementing AI-assisted operations for infrastructure management. Proven track record of reducing operational toil through engineering and automation.

  • Experience leading globally distributed teams supporting critical infrastructure. Ability to remain technically engaged while providing organizational leadership and strategy.

NVIDIA is widely considered to be one of the technology world's most desirable employers. We have some of the most intelligent and hardworking people in the world working for us. If you're creative and autonomous, we want to hear from you!

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 248,000 USD - 391,000 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 3, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Manager, Compute Core Engineering
Senior Manager, Compute Core Engineering

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 248,000 - 391,000
Equity
Benefits
Senior Manager, Compute Core Engineering
Senior Manager, Compute Core Engineering

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 248,000 - 391,000
Equity
Benefits
Senior Staff Site Reliability Engineer - Compute Core Engineering
Senior Staff Site Reliability Engineer - Compute Core Engineering

Socket.dev • California (MO)

Hybrid
USD 200,000 - 322,000
Equity
Benefits
Senior Staff Site Reliability Engineer - Compute Core Engineering
Senior Staff Site Reliability Engineer - Compute Core Engineering

NVIDIA • United States

On-site
USD 200,000 - 322,000
Senior Staff Platform Engineer
Senior Staff Platform Engineer

NVIDIA • Santa Clara (CA)

On-site
USD 200,000 - 322,000
Senior Staff Site Reliability Engineer - Compute Core Engineering
Senior Staff Site Reliability Engineer - Compute Core Engineering

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

On-site
USD 200,000 - 322,000
Senior Staff Platform Engineer
Senior Staff Platform Engineer

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 200,000 - 322,000
Senior Infrastructure Solutions Architect
Senior Infrastructure Solutions Architect

Nvidia Corporation in • Austin (TX)

On-site
USD 184,000 - 288,000
Equity
Benefits
Principal Software Engineer – Infrastructure
Principal Software Engineer – Infrastructure

NVIDIA • Redmond (WA)

On-site
USD 248,000 - 391,000
Principal Software Engineer - Compute Infrastructure
Principal Software Engineer - Compute Infrastructure

Nvidia Corporation in • Santa Clara (CA)

Hybrid
USD 248,000 - 391,000
Equity
Benefits