Senior Manager, Compute Core Engineering

NVIDIA AI

Santa Clara (CA)

Hybrid

USD 248,000 - 391,000

Full time

14 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

NVIDIA seeks a Senior Manager to lead the Compute Core Engineering team, directing strategy for core infrastructure services across data centers, cloud, and compute platforms. You will drive modernization, automation, and reliability while mentoring a global team of engineers.

You will shape the roadmap for scalable infrastructure, partnering with leadership and customers to deliver high-performance systems and AI-ready capabilities.

Qualifications

  • Bachelor's or Master's degree in a related technical field or equivalent experience.
  • 12+ years of experience in related fields with significant technical leadership.
  • Strong experience in infrastructure and distributed systems.
  • Solid knowledge of networking technologies and Linux.
  • Experience with container platforms and GitOps approaches.

Responsibilities

  • Lead and expand a global Compute Core Engineering team; recruit, mentor and develop talent.
  • Define strategy and roadmap for infrastructure services across data centers, cloud, and compute platforms.
  • Modernize legacy infrastructure into automated, self-service platforms with minimal disruption.
  • Provide technical leadership for DNS, DHCP, NTP/PTP, LDAP, and Linux infrastructure.
  • Establish engineering standards for performance, capacity, security, lifecycle management and disaster recovery.
  • Champion automation and AI-assisted operations to reduce manual work and improve reliability.
  • Lead the creation of automation platforms using Go, Python, Terraform, and GitOps.
  • Establish service health metrics and lead incident reviews and reliability programs.
  • Forecast demand and collaborate with teams to ensure scalable infrastructure.
  • Collaborate with leadership to build compelling infrastructure products.

Skills

Leadership
Distributed systems
Networking
Linux
Go
Python
GitOps
Automation

Education

Bachelor's or Master's degree in a related technical field or equivalent experience

Tools

Terraform
GitOps
Linux
DNS
LDAP
NTP/PTP

Job description

Job Requisition ID JR2024540

Job Category Engineering

Time Type Full time

We are seeking a Senior Manager - Compute Core Engineering to lead our Compute Core Engineering team. This opportunity is outstanding as it allows you to contribute to groundbreaking advancements in AI and computing technology, advancing NVIDIA's legacy of innovation forward. You will have the chance to work with dedicated individuals and be part of a world-class team committed to flawless execution and ambitious goals.

What You Will Be Doing
  • Lead and expand a group responsible for NVIDIA's global Compute Core Engineering Team. Recruit, mentor, and develop technical talent, encouraging a culture of ownership, collaboration, and innovation.
  • Define the strategy and roadmap for infrastructure services across NVIDIA's data centers, cloud environments, and compute platforms.
  • Drive the transformation of traditional services into automated, self-service platforms built for reliability and scalability. Modernize legacy infrastructure with minimal disruption.
  • Provide technical leadership for globally distributed services including DNS, DHCP, NTP/PTP, LDAP, and Linux infrastructure.
  • Establish engineering standards for performance, capacity, scalability, security, lifecycle management, and disaster recovery.
  • Build a culture centered around automation, operational excellence, and reducing manual work. Champion AI-assisted operations to automate diagnostics and routine infrastructure operations.
  • Partner with senior engineers to introduce technologies that improve infrastructure performance and resiliency. Lead the creation of automation platforms through Go, Python, Terraform, and GitOps approaches.
  • Establish metrics for service health, including availability, latency, capacity utilization, and infrastructure efficiency. Lead incident reviews and reliability improvement programs.
  • Lead forecasting activities and collaborate with various teams to ensure infrastructure can handle upcoming demands.
  • Collaborate closely with NVIDIA leadership and internal customers to build compelling infrastructure products.
What We Need To See
  • Bachelor's or Master's degree in a related technical field or equivalent experience.
  • 12+ overall years of experience in related fields, including significant technical leadership experience. 5+ years of engineering management experience.
  • Demonstrated experience crafting and managing critically important infrastructure.
  • Solid grasp of distributed systems architecture. Solid knowledge of networking technologies and protocols.
  • Deep experience with Linux operating systems. Experience with container platforms and technologies.
  • Solid experience in infrastructure automation and GitOps approaches. Software engineering experience with Go, Python, or similar languages.
  • Experience using SLIs, SLOs, and performance metrics.
  • Experience in capacity forecasting and lifecycle oversight.
  • Demonstrated capability to guide intricate modernization projects.
  • Strong communication skills to translate technical topics for senior leadership. Demonstrated capability to draw in, champion, and keep high-achieving teams.
Ways To Stand Out From The Crowd
  • Experience crafting and managing foundational infrastructure services. Experience supporting large AI, HPC, or accelerated computing environments.
  • Hands-on experience with advanced networking and infrastructure acceleration technologies. Deep comprehension of Linux kernel internals and improving system performance.
  • Experience modernizing legacy infrastructure into self-service platforms. Experience building engineering platforms around Terraform and GitOps or equivalent experience.
  • Experience implementing AI-assisted operations for infrastructure management. Proven track record of reducing operational toil through engineering and automation.
  • Experience leading globally distributed teams supporting critical infrastructure. Ability to remain technically engaged while providing organizational leadership and strategy.

NVIDIA is widely considered to be one of the technology world's most desirable employers. We have some of the most intelligent and hardworking people in the world working for us. If you're creative and autonomous, we want to hear from you!

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 248,000 USD - 391,000 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 3, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Manager, Compute Core Engineering
Senior Manager, Compute Core Engineering

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 248,000 - 391,000
Equity
Benefits
Senior Manager, Compute Core Engineering
Senior Manager, Compute Core Engineering

NVIDIA • Santa Clara (CA)

On-site
USD 248,000 - 391,000
Equity
Benefits
Senior Manager, Compute Core Engineering
Senior Manager, Compute Core Engineering

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 248,000 - 391,000
Equity
Benefits
Senior Staff Site Reliability Engineer - Compute Core Engineering
Senior Staff Site Reliability Engineer - Compute Core Engineering

NVIDIA • Santa Clara (CA)

On-site
USD 200,000 - 322,000
Equity
Benefits
Senior Staff Site Reliability Engineer - Compute Core Engineering
Senior Staff Site Reliability Engineer - Compute Core Engineering

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 200,000 - 322,000
Equity
Health benefits
Principal Software Engineer - Compute Infrastructure
Principal Software Engineer - Compute Infrastructure

Nvidia Corporation in • Santa Clara (CA)

Hybrid
USD 248,000 - 391,000
Equity
Benefits
Senior Staff Site Reliability Engineer - Compute Core Engineering
Senior Staff Site Reliability Engineer - Compute Core Engineering

Socket.dev • California (MO)

Hybrid
USD 200,000 - 322,000
Equity
Benefits
Senior Staff Platform Engineer
Senior Staff Platform Engineer

Nvidia Corporation • Santa Clara (CA)

On-site
USD 200,000 - 322,000
Equity
Comprehensive benefits
Senior Infrastructure Engineer - AI, Automation, Observability and Monitoring
Senior Infrastructure Engineer - AI, Automation, Observability and Monitoring

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 200,000 - 322,000
Senior Staff Site Reliability Engineer - Compute Core Engineering
Senior Staff Site Reliability Engineer - Compute Core Engineering

NVIDIA • United States

On-site
USD 200,000 - 322,000