Senior System Software Engineer, Performance - CUDA Driver

NVIDIA AI

Santa Clara (CA)

On-site

USD 184,000 - 357,000

Full time

16 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA seeks senior systems software engineers to accelerate the CUDA Driver, improving performance across CPUs, interconnects, and GPUs for the next generation of accelerated computing.

You will trace workloads, design features, and ship validated optimizations, collaborating across CUDA software, hardware, frameworks, and product teams. You will mentor engineers and influence CUDA software and GPU architectures.

Qualifications

  • BS/MS/PhD in Computer Science, Computer Engineering, Electrical Engineering, or equivalent with at least 7+ years of systems-software development experience.
  • Strong production C/C++ systems-programming experience in a complex codebase.
  • Solid operating systems and concurrency foundations including threads, synchronization, processes, virtual memory, and user/kernel interactions.
  • Strong computer architecture foundations including processors, memory hierarchy, caching and coherence, data movement, and system interconnects.
  • Proven track record improving real software performance through measurement, bottleneck identification, and profiling.
  • Sound technical judgment and ownership of ambiguous problems with clear cross-team communication.
  • Direct CUDA or GPU experience is valuable but not required if deep systems software foundations are present.

Responsibilities

  • Develop and ship performance-centric CUDA Driver features and programming-model capabilities from design through validation.
  • Diagnose complex performance problems through workload analysis, measurement and modeling, cross-layer root-cause isolation, and validation.
  • Optimize critical CUDA primitives, memory management and movement, CPU-GPU coordination, and interconnect paths for latency, throughput, and scalability.
  • Set performance goals for current and future platforms; characterize new platforms and drive readiness through releases.
  • Translate workload and platform evidence into CUDA API, programming model, system software, and future GPU architecture recommendations.
  • Set subsystem performance direction and mentor engineers tackling complex systems and performance challenges.
  • Influence technical decisions with evidence-based tradeoffs and strengthen implementations through design and code reviews.

Skills

C/C++ systems programming
Concurrency and OS fundamentals
Computer architecture foundations
Performance optimization
Root-cause analysis
Communication across teams
CUDA/GPU experience (optional)

Education

BS/MS/PhD in CS/CE/EE or equivalent

Job description

Job Requisition ID JR2009423 Job Category Engineering Time Type Full time We are looking for senior systems software engineers to make the CUDA Driver faster, more efficient, and ready for the next generation of accelerated computing. Our team develops performance-critical CUDA Driver features and systems software improvements that help AI, deep learning, HPC, and other CUDA-powered applications realize more of the performance available from NVIDIA GPUs!

Some of the hardest performance problems emerge not within one component, but at the boundaries among systems software, CPUs, interconnects, and GPUs. In this role, you will trace those problems from real workloads through the software and hardware stack, design and ship production features, and deliver validated optimizations for current and emerging platforms.

You will combine hands‑on engineering with broad technical influence, collaborating across CUDA software, hardware architecture, frameworks, applications, product, and customer‑facing teams. You will lead cross‑layer investigations, mentor engineers, and help set performance direction. Your work will help developers get more useful computing from NVIDIA GPUs today while helping shape the CUDA software and GPU architectures NVIDIA builds next.

What You'll Be Doing
  • Develop and ship performance‑centric CUDA Driver features and programming‑model capabilities from design through validation.
  • Diagnose complex performance problems through workload analysis, focused measurement and modeling, cross‑layer root‑cause isolation, and application‑level validation.
  • Optimize critical CUDA primitives, memory management and movement, CPU-GPU coordination, and interconnect paths for latency, throughput, bandwidth, efficiency, and scalability.
  • Establish performance goals for current and future platforms, characterize as new platforms come online, close software and hardware gaps, and drive performance readiness through releases.
  • Translate workload and platform evidence into CUDA API, programming model, system software, and future GPU architecture recommendations.
  • Set subsystem performance direction and mentor engineers tackling complex systems and performance challenges.
  • Influence technical decisions with clear performance evidence and tradeoffs and strengthen implementations through rigorous design and code reviews.
What We Need To See
  • A BS, MS, or PhD in Computer Science, Computer Engineering, Electrical Engineering, or a related field—or equivalent practical experience — with at least 7+ years of relevant systems‑software development experience.
  • Strong production C/C++ systems‑programming experience, including delivery of substantial features, optimizations, or production fixes in a complex codebase.
  • Strong operating systems and concurrency foundations, including threads, synchronization, processes, virtual memory, and user/kernel interactions.
  • Strong computer architecture foundations, including processors, memory hierarchy, caching and coherence, data movement, and system interconnects.
  • Demonstrated success improving real software performance: measuring behavior, identifying bottlenecks, implementing optimizations, and profiling to prove their effectiveness.
  • Sound technical judgment, ownership of ambiguous problems, and clear communication across organizational and disciplinary boundaries.
  • Direct CUDA or GPU experience is valuable but is not required when accompanied by deep systems software, operating systems, computer architecture, and performance‑engineering foundations.
Ways To Stand Out From The Crowd
  • Experience developing GPU or accelerator drivers, runtimes, kernel software, firmware, compilers, or other performance‑critical low‑level systems.
  • Experience with pre‑silicon analysis, platform bring‑up, performance modeling, or hardware/software co‑design.
  • Systems‑level performance experience with AI/DL, HPC, graphics, automotive, robotics, or similarly demanding workloads.
  • Evidence of technical inventions such as software‑performance patents, novel production designs, or measurement‑backed recommendations that influenced hardware revision or future architecture.
  • Python or another scripting language used for focused experimentation, data analysis, or visualization.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 6, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

System Software Engineer, Performance - CUDA Driver
System Software Engineer, Performance - CUDA Driver

Socket.dev • Santa Clara (CA)

On-site
USD 124,000 - 196,000
Equity
Benefits
System Software Engineer, Performance - CUDA Driver
System Software Engineer, Performance - CUDA Driver

NVIDIA • California (MO)

On-site
USD 124,000 - 196,000
Equity
Benefits
Senior System Software Engineer, Performance - CUDA Driver
Senior System Software Engineer, Performance - CUDA Driver

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 287,500
System Software Engineer, Performance – CUDA Driver
System Software Engineer, Performance – CUDA Driver

Nvidia • Jasper (AL)

On-site
USD 124,000 - 196,000
Equity
Benefits
System Software Engineer, Performance - CUDA Driver
System Software Engineer, Performance - CUDA Driver

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 124,000 - 196,000
System Software Engineer, Performance - CUDA Driver
System Software Engineer, Performance - CUDA Driver

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 124,000 - 196,000
Equity
Benefits
Senior Systems Software Engineer, CUDA Driver
Senior Systems Software Engineer, CUDA Driver

NVIDIA AI • Eugene (OR)

On-site
USD 184,000 - 288,000
Equity
Benefits
Senior System Software Engineer - CUDA Chips
Senior System Software Engineer - CUDA Chips

NVIDIA AI • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits package
Senior Software Engineer - CUDA Driver
Senior Software Engineer - CUDA Driver

NVIDIA AI • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior System Software Engineer - CUDA Chips
Senior System Software Engineer - CUDA Chips

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 152,000 - 288,000
Equity
Benefits