Senior Systems Software Engineer - Infrastructure

NVIDIA Gruppe

Santa Clara (CA)

On-site

USD 224,000 - 357,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NVIDIA is seeking a senior engineer to define the infrastructure and AI direction for its GPU Firmware Infrastructure team. You will shape the automation, tooling, and release pipelines that accelerate GPU innovation, working across firmware, hardware, and software boundaries.

We value deep expertise in CI/CD, Python, and distributed systems, plus a track record of mentoring engineers and delivering scalable platform solutions. Join a diverse, top-talent team building AI supercomputers.

Qualifications

  • BS or MS in EE/CS/CE (or equivalent experience).
  • 12+ years of infrastructure, platform, DevOps, or SRE engineering.
  • Deep expertise in modern CI/CD and test automation architecture: pipeline design, build system internals, caching, testing frameworks, and failure modes at scale.
  • Strong Python fluency, infrastructure-as-code experience, and scripting ability.
  • Experience with distributed systems, cloud or on-prem compute fleets, containers, and orchestration (K8s or equivalent).
  • Strong understanding of database concepts, schema design, object modeling, SQL or non-SQL.
  • Proven ownership of cross-org production systems and lessons learned from supporting them.
  • Excellent communication, leadership, and mentoring skills.

Responsibilities

  • Set technical direction for major platform areas, leading architecture, roadmap, and long-term health across multiple product cycles.
  • Design and lead infrastructure initiatives that span team boundaries, aligning with firmware, hardware, and business partners.
  • Own platform reliability: define SLIs/SLOs, instrument gaps, reduce MTTR, automate incident response.
  • Identify automation opportunities and deliver.
  • Architect and invent AI-powered workflows that boost productivity across the team.
  • Set technical standards through design reviews, documentation, and mentoring to improve decision quality.

Skills

Python
CI/CD
Test automation
Distributed systems
Mentoring
Communication

Education

BS/MS in EE/CS/CE

Tools

Kubernetes
Infrastructure as Code

Job description

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.

Do you derive more satisfaction from eliminating a manual process than from completing one? Do you look at a build system, a release pipeline, or a testing fleet and immediately see the three things held together with duct-tape, and the one architectural decision that would make all three unnecessary? We’re looking for a senior engineer to invent, set, and construct the infrastructure and AI direction that enables us to create the software and firmware that drive the world’s best GPUs. Join us, the GPU Firmware Infrastructure team, whose mission is crafting extraordinary automation, tools, and practices that accelerate NVIDIA’s GPU innovation. We’re responsible for the substrate: dev tools, build compute, test farms, artifact and release pipelines, secure signing services, and the observability that proves it all works. When our platforms are fast and trustworthy, every product ships faster! This is your chance to make waves in the industry while working alongside some of the most top-valued diverse set of minds in the business, building the best AI Supercomputers.

What you’ll be doing:
  • Set technical direction for major platform areas, leading the architecture, roadmap, and long-term health across multiple product cycles
  • Design and lead infrastructure initiatives that span team boundaries, building alignment with firmware, hardware, and business partners
  • Own platform reliability: define SLIs and SLOs, instrument what isn’t measured today, cut MTTR, and automate the incident response when the platform breaks
  • Find the automation opportunities nobody has asked for yet, make the case for it and deliver
  • Architect and invent AI-powered workflows that multiply everyone’s output
  • Set the technical bar through design review, technical writing, and mentoring, so the team makes better decisions without you in the room
What we need to see:
  • BS or MS degree in EE/CS/CE (or equivalent experience)
  • 12+ years of infrastructure, platform, DevOps, or SRE engineering
  • Deep expertise in modern CI/CD and test automation architecture: pipeline design, build system internals, caching, testing frameworks, and the failure modes for architectures that only appear at scale
  • Strong Python fluency, infrastructure-as-code experience, and scripting ability
  • History in distributed systems, cloud or on-prem compute fleets, containers, and orchestration (K8s or equivalent)
  • Strong understanding of database concepts, schema design, object modeling, SQL or non-SQL
  • Past ownership of cross-org production system(s) and all the lessons you’ve learned from supporting them
  • Outstanding communication, leading, and teaching skills: the ability to articulate WHY a solution is correct, write clear requirements, review critically, and guide others to understanding
  • History of growing other engineers, formally or informally
Ways to stand out from the crowd:
  • Sense of humor heavily encouraged, but not required
  • Track record of writing technical proposals, design docs, or architecture decisions that others have acted on independently
  • Experience using AI-assisted development tools: knowing when to trust, when to verify, and when to throw away the output
  • Experience in developing device BIOS, firmware, or other low-level embedded software
  • Pride and passion evident in your work
Your base salary will be determined based on your location, experience, and the pay of employees in similar positions.

The base salary range is 224,000 USD - 356,500 USD. You will also be eligible for equity and benefits.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Systems Software Engineer - Infrastructure
Senior Systems Software Engineer - Infrastructure

Socket.dev • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Senior Systems Software Engineer - Infrastructure
Senior Systems Software Engineer - Infrastructure

NVIDIA AI • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Equity
Benefits
Senior Systems Software Engineer - Infrastructure
Senior Systems Software Engineer - Infrastructure

NVIDIA • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Equity
Benefits
Systems Software Engineer - Infrastructure
Systems Software Engineer - Infrastructure

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 152,000 - 242,000
Equity
Comprehensive benefits
Senior Systems Software Engineer - Infrastructure
Senior Systems Software Engineer - Infrastructure

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Equity
Benefits
Systems Software Engineer - Infrastructure
Systems Software Engineer - Infrastructure

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 152,000 - 242,000
Equity
Benefits
Senior Staff Platform Engineer
Senior Staff Platform Engineer

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 200,000 - 322,000
Distinguished Engineer, System Software Integration
Distinguished Engineer, System Software Integration

Nvidia Corporation • Santa Clara (CA)

On-site
USD 320,000 - 489,000
Equity
Benefits package
Distinguished Engineer, System Software Integration
Distinguished Engineer, System Software Integration

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 320,000 - 489,000
Equity compensation
Benefits package
Competitive salary
Distinguished Engineer, System Software Integration
Distinguished Engineer, System Software Integration

NVIDIA Gruppe • California (MO)

On-site
USD 320,000 - 489,000
Equity
Benefits package