Senior Manager, Test, Manufacturability, Reliability and Quality

Nvidia Corporation

Santa Clara (CA)

On-site

USD 232,000 - 368,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Nvidia Corporation in Santa Clara is looking for a Senior Manager to lead the Test, Manufacturability, Reliability & Quality organization. This position is critical in ensuring products are built at scale and trusted in the field. The ideal candidate has over 12 years of experience in production test and reliability engineering, with exceptional leadership skills to develop high-performance teams.

The base salary is between $232,000 and $368,000, along with equity and additional benefits.

Qualifications

  • 12+ years in production test, system validation, or reliability engineering in a high-volume silicon environment.
  • 5+ years in management with a strong foundation in engineering principles.
  • Hands-on experience with production test stacks and methodologies.

Responsibilities

  • Lead technical execution and drive resolution of complex failures in multi-functional projects.
  • Provide decision-ready options from raw data for executive leadership.
  • Hire and develop a high-performing team across multiple locations.

Skills

Production test expertise
Reliability engineering
Team leadership
Analytical skills

Education

BS, MS, or PhD in Electrical Engineering, Computer Engineering, Systems Engineering

Tools

ATE
Control Run methodology
FinFET-class manufacturing processes

Job description

NVIDIA's Silicon Co-Design Group is the team that gets every GPU, SoC, and CPU silicon program from first power-on to high-volume production. We are hiring a Senior Manager to lead our Test, Manufacturability, Reliability & Quality (TMRQ) organization. This is not a coordination role. Your work decides if a product can be built at scale and trusted in the field. These include production test development (SLT, BLT), control run flow, system reliability stress (HTOL), platform- and board-level manufacturing issue closure, and field diagnostic test development. You lead a team of individual contributors and a first-line manager at the layer where silicon, platform, and software collide with manufacturing reality. Decisions you make show up in yield curves, production ramp, and customer escapes. You are the leader the program turns to when a build is stuck, a control run is fallout-heavy, or a field return points back at silicon. The exceptional hire also uses AI deliberately - with demonstrated workflow impact and the judgment to know where it compresses real work and where it introduces risk.

What you will be doing:
  • Keep programs moving. Own the technical execution and velocity of SLT, BLT, Board/Chip/Rack CR, and system reliability stress (HTOL) across every GPU, SoC, and CPU silicon program.
  • Close the hardest multi-functional failures. Resolve Vmin and binning escapes, performance shortfalls, and power anomalies by driving root-cause across design, methodology, DFT, ATE, package, software/firmware, and manufacturing - and own the WARs and productized fixes through to confirmation.
  • Give leadership the clarity to act. Convert raw integration signals - CR fallout, BLT/SLT yield, SHTOL/CHTOL data, RMA trends, customer escalations - into decision-ready options that enable executive leadership to act with confidence on POR, QS/PS gates, and ramp-impacting risks.
  • Raise the execution bar across every program. Establish operational objectives, and refine closed-loop methodologies that improve cycle time, first-pass yield, test-time reduction, escape rate, and DPPM.
  • Prevent field escapes from recurring. Take ownership of manufacturing issues at both platform and board levels. Conduct root-cause analysis on returns and field escapes. Use field signals to develop screening, process, and build changes that improve quality.
  • Build the team the organization depends on. Hire, develop, and retain your manager and lead engineers across multiple geographies - grow ICs into recognized subsystem experts and develop your first-line managers into independently operating leaders.
What we need to see:
  • BS, MS, or PhD in Electrical Engineering, Computer Engineering, Systems Engineering, or related field (or equivalent experience).
  • 12+ years in production test, system validation, post-silicon integration, or reliability engineering in a high-volume silicon environment, including 5+ years in management.
  • Strong EE fundamentals across DFT, BIST, digital design, computer architecture, fault models and fault analysis, power and timing analysis, sampling, and statistics.
  • Hands-on expertise across the production test stack (ATE, SLT, BLT), Control Run methodology, FinFET-class manufacturing processes, and PVT/binning dependencies.
Ways to stand out from the crowd:
  • Background in GPU, CPU, AI accelerator, or other large-SoC programs - you know how complexity at scale changes the failure landscape.
  • Applied AI tools in production test or debug workflows - faster analysis, smarter triage, automated reporting - and can describe the outcome and the guardrails you put in place.
  • Track record of building and retaining senior technical talent in a highly matrixed organization, including guiding first-line managers to become autonomous leaders.
  • Ability to operate independently on hard, ambiguous problems - and collaborate clearly across functions.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 232,000 USD - 368,000 USD. You will also be eligible for equity and benefits.

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior System Level Testability Engineer
Senior System Level Testability Engineer

NVIDIA Corporation • Santa Clara (CA)

Hybrid
USD 168,000 - 311,000
Equity
Benefits
Director, System Level Testing - Product Development Engineering
Director, System Level Testing - Product Development Engineering

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 272,000 - 426,000
Equity
Benefits
Director, System Level Testing - Product Development Engineering
Director, System Level Testing - Product Development Engineering

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 272,000 - 426,000
Equity
Benefits
Director, System Level Testing - Product Development Engineering
Director, System Level Testing - Product Development Engineering

NVIDIA • Santa Clara (CA)

On-site
USD 272,000 - 426,000
Director, System Level Testing - Product Development Engineering
Director, System Level Testing - Product Development Engineering

NVIDIA AI • Santa Clara (CA)

On-site
USD 272,000 - 426,000
Equity
Benefits
Manager, System Test Engineering
Manager, System Test Engineering

Nvidia Corporation • Santa Clara (CA)

On-site
USD 152,000 - 253,000
Post-Silicon Validation and Methodology Engineer
Post-Silicon Validation and Methodology Engineer

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 136,000 - 219,000
Senior Board Test Engineer
Senior Board Test Engineer

Nvidia Corporation • Santa Clara (CA)

On-site
USD 132,000 - 207,000
Equity options
Generous benefits package
Senior System Test Engineer
Senior System Test Engineer

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 160,000 - 270,000
Senior System Test Engineer
Senior System Test Engineer

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

On-site
USD 152,000 - 288,000