Manager, System Software Engineering - Factory

NVIDIA AI

Santa Clara (CA)

On-site

USD 224,000 - 356,500

Full time

5 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA in Santa Clara, CA seeks an Engineering Manager to lead Factory System Software and Diagnostics Integration, delivering embedded code and updates for GPU/DPU products across global factory partners. You will shape end-to-end workflows, ensure high-quality releases land on factory floors, and partner with architects, firmware, SWQA, product management, and ODM/OEM teams.

The role demands 10+ years in software with system/firmware focus, 3+ years of engineering leadership, and strong

Qualifications

  • 10+ years in software with specialization in system software or firmware.
  • 3+ years of engineering management or technical leadership with distributed teams.
  • BS, MS, or PhD in CS, CE, EE, or related field (or equivalent).
  • Proven track record of shipping scalable server products through factory ramps with ODM/OEM partners.
  • Strong written and oral communication, including executive-level reporting, and teamwork.
  • Experience across cross-functional collaboration with hardware, firmware, diagnostics and QA.
  • Willingness to work across time zones and travel as needed.

Responsibilities

  • Build, lead, mentor, and grow a global factory engineering team spanning the US and Taiwan.
  • Define Factory readiness scope and workflows for rack-scale products with PM, architects and program management.
  • Own technical leadership for firmware, software, and diagnostics releases reaching factories.
  • Left-shift release quality with CI/CD and quality gates, reporting progress to stakeholders.
  • Own the factory escalation path: SLAs, 24x7 coverage, root-cause analysis, and burn-down.
  • Shape the team's roadmap and drive automation and AI-assisted validation and triage.
  • Continuously analyze processes to identify improvements and publish SOPs.

Skills

C/C++
Python
Firmware
System software
Leadership
Distributed teams
Executive reporting
Automation
Diagnostics
Hardware collaboration

Education

BS/MS/PhD in CS/CE/EE

Tools

Git
Perforce
Jira

Job description

Job Requisition ID

JR2021001

Job Category

Engineering

Time Type

Full time

NVIDIA’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern deep learning — the next era of computing — with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as “the AI computing company.” We\'re looking to grow our company and establish teams with the most thoughtful people in the world.

We are the Datacenter System Software team, and we are looking for a highly motivated, creative Engineering Manager to drive Factory System Software and Diagnostics Integration end to end. You will build and lead a global engineering team delivering embedded code, application programs, and diagnostic updates. These updates support factories building NVIDIA\'s GPU- and DPU-based products. This includes tightly coupled rack-scale systems such as GB200/GB300 NVL72 and next-generation platforms. The work covers concurrent NPI ramps and sustaining production. You will partner with system architects, firmware developers, SWQA, product engineering, compliance and security teams, program and product management, and ODM/CM manufacturing partners to ensure the highest-quality releases land on factory floors — and that no bug is discovered there first. Join us at the forefront of technological advancement.

What You’ll Be Doing
  • Build, lead, mentor, and grow a global factory engineering team spanning the US and Taiwan — operating a follow-the-sun coverage model with on-site presence at ODM/CM partner factories. Own hiring, career development, calibration, and succession planning.
  • Define Factory readiness scope and workflows for rack scale products coordinating multi-functionally with product management, technical architects and program management. Deliver those workflows through the validation matrix, ensuring delivered firmware and software is of the highest quality. Solutions must scale and be resilient.
  • Own technical leadership for how firmware, software, and diagnostics releases reach factories building rack-scale systems. These systems include tightly coupled compute and switch trays. Build the end-to-end infrastructure and workflows that ensure every release arrives with efficient quality.
  • Left-shift release quality: partner with all matrixed organizations — developers, SWQA, and product engineering — in a fast-moving environment with end-to-end CI/CD so that no bug is first found at a factory site. Enforce well-placed quality gates at every product landmark, publish and track indicators at a regular cadence, and report release progress to collaborators and executives.
  • Own the factory escalation path: triage SLAs, 24×7 coverage, failure root-cause and deflection, and bonepile burn-down — minimizing line-down time through NPI ramps and mass production.
  • Shape the team\'s roadmap and drive innovation with a strong focus on automation and AI-assisted validation and triage — automating station readiness, firmware-update flows, and log triage so senior engineering time shifts from setup to analysis.
  • Continuously analyze factory processes, systems, and workflows to identify improvement and optimization opportunities; remove bottlenecks, document and publish standard operating procedures (SOPs), and ensure the team performs in the most efficient and transparent way against measurable targets.

What We Need To See

  • 10+ overall years in the software industry with specialization in system software and/or firmware development.
  • 3+ years of engineering management or technical leadership experience, including building and leading geographically distributed teams.
  • BS, MS, or PhD in CS, CE, EE, or a related technical field — or equivalent experience.
  • Proven track record of shipping scalable server products through factory ramps — from NPI bring-up to mass production — collaborating with hardware, firmware, manufacturing, diagnostics, and QA teams.
  • Experience working with ODM/OEM partners to deliver quality servers and solutions for large-scale data centers.
  • A self-starter who loves finding creative solutions to complicated problems, with excellent written and oral communication skills — including executive-level reporting — strong work ethic, and dedication to teamwork.
  • Flexibility to work and communicate effectively across teams, partners, and time zones.
Ways To Stand Out From The Crowd
  • Experience leading bring-up for sophisticated rack-scale compute architectures like GB200/GB300 NVL72.
  • Familiarity with manufacturing test flows (L6/L10/L11/L12 stations), factory test coverage, and MES integration.
  • Hands-on experience with x86/ARM system architecture and coding (C/C++, Python). Experience with SCM (Git, Perforce) and project management tools (Jira).
  • Track record of integrating AI/LLM tooling into engineering workflows — for triage, validation, log analysis, or test generation.
  • Experience standing up follow-the-sun support organizations with measurable response SLAs.

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you\'re creative and autonomous, we want to hear from you!

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 224,000 USD - 356,500 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 23, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Global Factory Systems Engineering Manager - Diagnostics
Global Factory Systems Engineering Manager - Diagnostics

NVIDIA • Santa Clara (CA)

On-site
USD 224,000 - 356,500
Manager, System Software Engineering - Factory
Manager, System Software Engineering - Factory

NVIDIA • Santa Clara (CA)

On-site
USD 224,000 - 356,500
Manager, System Software Engineering - Factory
Manager, System Software Engineering - Factory

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Senior Software Engineer - Manufacturing and Factory
Senior Software Engineer - Manufacturing and Factory

NVIDIA • Austin (TX)

On-site
USD 184,000 - 288,000
Equity
Comprehensive benefits package
Senior Software Engineer - Manufacturing and Factory
Senior Software Engineer - Manufacturing and Factory

NVIDIA • Durham (NC)

On-site
USD 184,000 - 357,000
Equity
Comprehensive benefits package
Senior Software Engineer - Manufacturing and Factory
Senior Software Engineer - Manufacturing and Factory

NVIDIA • Redmond (WA)

On-site
USD 184,000 - 357,000
Equity
Benefits package
Rack-Scale Systems Engineer: Equity & Impact
Rack-Scale Systems Engineer: Equity & Impact

NVIDIA • Austin (TX)

On-site
USD 184,000 - 288,000
Equity
Comprehensive benefits package
Senior Director, Board Product Development Engineering
Senior Director, Board Product Development Engineering

NVIDIA • Santa Clara (CA)

On-site
USD 332,000 - 500,000
Equity
Benefits
International travel
Technical Program Manager, New Product Deployment Readiness
Technical Program Manager, New Product Deployment Readiness

NVIDIA Corporation • Town of Texas (WI), Northern (KY)

Hybrid
USD 168,000 - 322,000
Equity
Benefits
Senior Systems Software Engineer - Infrastructure
Senior Systems Software Engineer - Infrastructure

NVIDIA AI • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Equity
Benefits