Engineering Manager, Reliability Engineering Flywheel – EDA Infrastructure

NVIDIA

Massachusetts

On-site

USD 260,000 - 431,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA’s EDA Infrastructure organization seeks an engineering manager to lead the team responsible for operational processes and platforms across incident management, maintenance, on‑call, issue management, and customer‑serving readiness.

You will own the roadmap and delivery, from defining how teams work to building the tools they use, partnering with infrastructure and service owners to improve reliability and reduce manual work.

Qualifications

  • 10+ years of software engineering or related experience.
  • 5+ years of engineering leadership managing teams or complex programs.
  • Knowledge of operational processes and supporting platforms.

Responsibilities

  • Lead a team and own the roadmap for operational processes and platforms, from requirements and delivery through adoption and results.
  • Set technical direction, priorities, and guide execution across engineering and operational disciplines.
  • Partner with infrastructure, product, and security to establish consistent practices for incident response, maintenance, on-call, issue management, and customer-ready readiness.
  • Hire and develop engineers and technical leads, building a team with clear ownership and accountability.
  • Align priorities across teams, communicate progress and risks, and provide technical leadership during major incidents.

Skills

Software engineering
Engineering leadership
Operational processes
Communication

Education

BS degree or equivalent

Job description

NVIDIA’s EDA Infrastructure organization builds and operates the systems that support chip development. We are looking for an engineering manager to lead the team responsible for operational processes and platforms across incident management, maintenance, on‑call, issue management, and customer‑serving readiness.

You will own the roadmap and delivery, from defining how teams work to building the tools they use. You will partner with infrastructure and service owners to improve reliability, reduce manual work, and ensure services are ready to support customers. Your team will use automation, AI, and lessons from operational events to drive improvements.

What You’ll Be Doing
  • Lead a team and own the roadmap for operational processes and platforms, from requirements and delivery through adoption and results.
  • Set technical direction, prioritize work, and guide execution across engineering and operational disciplines.
  • Partner with infrastructure, product, and security teams to establish consistent practices for incident response, maintenance, on‑call, issue management, and customer‑serving readiness.
  • Hire and develop engineers and technical leads, building a team with clear ownership and accountability.
  • Align priorities across teams, communicate progress and risks, and provide technical leadership during major incidents.
What We Need To See
  • BS degree or equivalent experience with 10+ overall years of software engineering or related experience, including 5+ years of engineering leadership managing teams or complex technical programs.
  • Knowledge of operational processes and supporting platforms, including roadmap, delivery, adoption, and improvement.
  • Strong technical judgment in software architecture, platform integration, and engineering tradeoffs.
  • Clear communication with engineers, cross‑functional partners and executive stakeholders.
  • A record of developing engineers, growing teams, and delivering results under pressure.
Ways To Stand Out From The Crowd
  • Established readiness standards covering service ownership, support coverage, and reliability objectives.
  • Built, integrated, and scaled platforms pertaining the incident, maintenance, customer experience management, along with on‑call and production readiness
  • Applied AI or LLMs to improve triage, knowledge retrieval, incident analysis, or automation.
  • Supported EDA, large‑scale compute, or hybrid infrastructure with complex dependencies and demanding availability requirements.

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward‑thinking and hard‑working people in the world working for us. Are you creative and autonomous?

Do you love a challenge? If so, we want to hear from you. NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High‑Performance Computing and Visualization. The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our products and services. If you’re creative and self‑motivated, we want to hear from you! NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High‑Performance Computing and Visualization. The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our products and services.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 224,000 USD – 356,500 USD for Level 3, and 272,000 USD – 431,250 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 26, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Engineering Manager, Reliability Engineering Flywheel – EDA Infrastructure
Engineering Manager, Reliability Engineering Flywheel – EDA Infrastructure

NVIDIA • California (MO)

On-site
USD 224,000 - 431,000
Equity
Benefits
Engineering Manager, Reliability Engineering Flywheel – EDA Infrastructure
Engineering Manager, Reliability Engineering Flywheel – EDA Infrastructure

NVIDIA • Washington

On-site
USD 224,000 - 431,000
Equity
Benefits
Engineering Manager, Reliability Engineering Flywheel - EDA Infrastructure
Engineering Manager, Reliability Engineering Flywheel - EDA Infrastructure

NVIDIA • Redmond (WA)

On-site
USD 270,000 - 431,000
Engineering Manager, Reliability Engineering Flywheel - EDA Infrastructure
Engineering Manager, Reliability Engineering Flywheel - EDA Infrastructure

NVIDIA • Washington

On-site
USD 250,000 - 420,000
Equity
Benefits
Engineering Manager, Reliability Engineering Flywheel - EDA Infrastructure
Engineering Manager, Reliability Engineering Flywheel - EDA Infrastructure

NVIDIA Corporation • Redmond (WA)

On-site
USD 224,000 - 431,000
Equity
Benefits
Engineering Manager, Reliability Engineering Flywheel - EDA Infrastructure
Engineering Manager, Reliability Engineering Flywheel - EDA Infrastructure

NVIDIA Corporation • Northern (KY)

Hybrid
USD 224,000 - 431,000
Software Engineer, System Validation - EDA Infrastructure
Software Engineer, System Validation - EDA Infrastructure

NVIDIA • Austin (TX)

On-site
USD 224,000 - 431,000
Equity
Benefits
Senior Systems Software Engineer- EDA Infrastructure
Senior Systems Software Engineer- EDA Infrastructure

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior Systems Software Engineer- EDA Infrastructure
Senior Systems Software Engineer- EDA Infrastructure

NVIDIA Gruppe • California (MO)

On-site
USD 184,000 - 357,000
Equity
Benefits
Head of Global Customer Engineering - Internal EDA Infrastructure
Head of Global Customer Engineering - Internal EDA Infrastructure

NVIDIA • Seattle (WA)

On-site
USD 272,000 - 489,000