Principal Systems Software Engineer, LPU

NVIDIA Gruppe

Santa Clara (CA)

On-site

USD 272,000 - 431,250

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NVIDIA Gruppe is seeking a Principal Software Engineer for its LPX System Software team in Santa Clara, California. This role involves building foundational software that optimizes complex compute architectures. You will design drivers and manage workloads, influencing both architecture and API contracts.

A master's degree in a relevant STEM field and over 12 years of experience are required, with deep expertise in Rust programming necessary. The salary ranges between $272,000 and $431,250, plus equity and benefits.

Qualifications

  • 12+ years building production system software.
  • Deep systems-programming expertise with Rust at the hardware or kernel boundary.
  • Experience with API compatibility and multi-repository codebases.

Responsibilities

  • Shape architecture of hardware abstraction layers.
  • Design and implement drivers and runtimes.
  • Lead new platform bring-up and NPI for new boards.

Skills

Rust programming
Systems programming
API design
Linux driver experience
Distributed systems

Education

MS in CS, CE, EE, or related STEM field

Job description

We are now looking for a Principal Software Engineer for LPX System Software! NVIDIA’s LPX System Software team builds the foundational software that turns a novel deterministic compute architecture into a platform that compiler teams and data center operators can rely on. We shift complexity out of silicon and into software: the hardware abstraction layers, core system libraries, drivers, and runtime components that workloads enter the platform through. We build this stack in Rust. For system software living at the boundary between hardware and everything above it, we treat memory safety, explicit ownership, and long-lived API stability as the baseline rather than the goal — the foundation that lets us spend our judgment on the hard problems instead of on classes of bugs that should not exist.

What you’ll be doing:
  • Shape the architecture of the hardware abstraction layers and core system libraries, and own the API contracts for the components you lead.
  • Design and implement drivers, runtimes, and data movement and aggregation pipelines that execute workloads on novel silicon.
  • Build runtime interfaces for launching, monitoring, and managing workloads at production scale.
  • Drive triage of the most difficult sequencing, initialization, and cross‑component runtime failures, and produce root‑cause analyses that change how the system is built.
  • Lead new platform bring‑up and NPI for new boards and silicon, in tight partnership with hardware engineering, compiler teams, and data center operations.
  • Multiply the team — establish the agent‑assisted engineering practices, reusable abstractions, diagnostics, and documentation that let everyone move faster without destabilizing the platform.
  • Communicate architecture and design tradeoffs clearly, in writing and in diagrams, to audiences ranging from individual engineers to executive staff.
What we need to see:
  • MS in CS, CE, EE, or a related STEM field, or equivalent experience, and 12+ years building production system software.
  • Deep systems‑programming expertise, with Rust as your language of choice for low-level work. You have shipped production Rust at the hardware or kernel boundary — drivers, firmware, runtimes, or similar — and you can articulate from experience where Rust earns its keep in system software and where it costs you.
  • A track record of designing and evolving libraries and APIs meant to be supported for years, including ABI and compatibility discipline.
  • Fluency in large, multi‑repository codebases with layered dependencies.
  • Demonstrated leadership driving triage of difficult reliability issues to clear, written root‑cause analysis.
  • Low‑level platform experience: firmware and boot flows, RTOS, BMCs/MCUs, RISC‑V, or closely related system software.
  • Linux driver or kernel‑adjacent experience (for example, VFIO or similar subsystems).
  • Hardware bring‑up and system triage experience: fault analysis, diagnostics, and validation in lab environments.
  • An established habit: building with AI coding agents — not as a novelty, but as a way you already ship and raise leverage. You can speak to how you design work to be agent‑amenable and where you keep humans in the loop.
Ways to stand out from the crowd:
  • Experience having built Rust system software at the scale of a hyperscaler or a Rust‑native hardware company.
  • Distributed systems experience: gRPC and RPC frameworks, coordination and telemetry patterns, MPI. Inference systems and token serving experience (vLLM or similar serving and runtime stacks) a huge plus.
  • Experience shipping and supporting customer‑facing SDKs, including documentation and ABI compatibility practices.
  • Production readiness and delivery depth: CI/CD and release workflows, monitoring and alerting practices, Kubernetes, and data center operational workflows.
Compensation

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 272,000 USD - 431,250 USD. You will also be eligible for equity and benefits.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Software Engineer - Rack Scale Systems Infrastructure
Principal Software Engineer - Rack Scale Systems Infrastructure

2100 NVIDIA USA • Santa Clara (CA)

On-site
USD 272,000 - 431,250
Senior Hardware Systems Engineer - LPU Platform Pathfinding
Senior Hardware Systems Engineer - LPU Platform Pathfinding

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Competitive salaries
Comprehensive benefits package
Equity
Senior Hardware Systems Engineer - LPU Platform Pathfinding
Senior Hardware Systems Engineer - LPU Platform Pathfinding

Segment (Twilio) • Santa Clara (CA)

On-site
USD 184,000 - 287,500
Principal Software Engineer, Rack-Scale System Software — CSP Engagements
Principal Software Engineer, Rack-Scale System Software — CSP Engagements

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 272,000 - 431,000
Equity
Benefits
Principal Software Engineer – CSP Engagements
Principal Software Engineer – CSP Engagements

NVIDIA • Santa Clara (CA)

On-site
USD 272,000 - 431,250
Equity
Benefits package
Senior Hardware Systems Architect - LPU
Senior Hardware Systems Architect - LPU

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 356,500
Equity
Benefits package
Principal Software Engineer, Rack-Scale System Software - CSP Engagements
Principal Software Engineer, Rack-Scale System Software - CSP Engagements

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 272,000 - 432,000
Equity
Comprehensive benefits package
Principal Software Engineer
Principal Software Engineer

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 272,000 - 431,000
Equity
Benefits
Senior Systems Software Engineer - Infrastructure
Senior Systems Software Engineer - Infrastructure

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Principal Software Engineer
Principal Software Engineer

NVIDIA • Hillsboro (OR)

On-site
USD 272,000 - 431,000
Equity
Benefits