Senior Performance Compiler Engineer - Triton

NVIDIA

California (MO)

On-site

USD 184,000 - 287,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Generous benefits package

Job summary

NVIDIA is seeking a Senior Performance Compiler Engineer in California, USA to work on the Triton compiler project. The role involves optimizing AI performance on NVIDIA GPUs using new technologies and compilers. Candidates should have over 8 years of experience and strong C++ skills, along with knowledge in parallel programming and computer architecture. NVIDIA offers a competitive salary range of $184,000 to $287,500 along with equity and benefits, and is committed to fostering a diverse work environment.

Qualifications

  • 8+ years of relevant industry experience in software development.
  • Demonstrated strong C++ programming and software design skills.
  • Solid understanding of computer architecture and assembly-level programming.

Responsibilities

  • Investigate NVIDIA GPU hardware architecture and programming models.
  • Design and implement compiler technology using MLIR.
  • Collaborate with hardware architects and the CUDA compiler team.

Skills

C++ programming
Software design
Performance analysis
Parallel programming
CUDA/OpenCL programming

Education

Bachelor, Masters or Ph.D. in Computer Science or related field

Tools

CUDA
OpenCL
MLIR

Job description

NVIDIA's invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern AI – the next era of computing – with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as “the AI computing company”.

We’re looking for a Senior Performance Compiler Engineer to join our team and work on the open‑source Triton compiler project. This opportunity involves working with new technologies and using compilers to improve AI performance on NVIDIA GPUs. Your work will enable breakthroughs in large language models, agents, and other high‑impact AI applications, accelerating both training and inference. You will be immersed in a diverse, supportive environment where everyone is inspired to do their best work, pushing the limits of what’s possible.

What You’ll Be Doing
  • Investigating the latest and future NVIDIA GPU hardware architecture and programming models.
  • Working on the frontier of AI by understanding advanced algorithms (like attention sinks and MoEs) and numerics (like block‑scaled floating point) to identify new opportunities for optimization.
  • Designing and implementing compiler technology using MLIR to optimize high‑level kernel descriptions (written in Triton’s Python DSL), with a focus on generating efficient, low‑level GPU code. When vital, you’ll also be able to use inline PTX to hand‑tune critical code paths and extract peak performance from the hardware.
  • Engaging in a dynamic, iterative process of optimization—sometimes starting with the kernel, sometimes with the compiler—to find the most efficient path to peak performance.
  • Collaborating with teams across NVIDIA, including hardware architects and the CUDA compiler team, to influence future products and ensure we are always operating at maximum efficiency.
What We Need To See
  • Bachelor, Masters or Ph.D. degree or equivalent experience in Computer Science, Computer Engineering, Applied Math, or a related field.
  • 8+ years of relevant industry experience in software development.
  • Demonstrated strong C++ programming and software design skills, with an emphasis on performance analysis and debugging.
  • Experience in parallel programming, including CUDA/OpenCL GPU programming or other parallel models such as OpenMP.
  • Solid understanding of computer architecture and hands‑on experience with assembly‑level programming.
Ways To Stand Out From The Crowd
  • Experience in tuning BLAS or deep learning library kernels.
  • Background in numerics and linear algebra.
  • Experience with machine learning compilers like TVM or MLIR.
  • Contributions to open-source projects, especially in the AI/ML or compiler space.
  • Familiarity with the latest research in AI algorithms and numerics as well as a strong track record of contributions to open‑source projects, particularly in the AI/ML, compiler, or high‑performance computing domains.

With competitive salaries and a generous benefits package, we are widely considered to be one of the technology world’s most desirable employers. We have some of the most forward‑thinking and hardworking people in the world working for us and, due to unprecedented growth, our exclusive engineering teams are rapidly expanding. If you’re a creative and autonomous engineer with a real passion for technology, we want to hear from you.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD – 287,500 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until May 12, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering a diverse work environment and proud to be an equal‑opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Performance Compiler Engineer - Triton
Senior Performance Compiler Engineer - Triton

NVIDIA • United States

On-site
USD 184,000 - 288,000
Senior Performance Compiler Engineer - Triton
Senior Performance Compiler Engineer - Triton

NVIDIA • Town of Texas (WI)

On-site
USD 184,000 - 288,000
Competitive salary
Equity
Generous benefits package
Senior Performance Compiler Engineer - Triton
Senior Performance Compiler Engineer - Triton

NVIDIA • Redmond (WA)

On-site
USD 184,000 - 288,000
Equity
Generous benefits package
Senior Performance Compiler Engineer - Triton
Senior Performance Compiler Engineer - Triton

NVIDIA Corporation • Redmond (WA)

On-site
USD 184,000 - 288,000
Generous benefits package
Equity options
Competitive salaries
Senior AI Compiler Engineer, Algorithms and Code-Generation
Senior AI Compiler Engineer, Algorithms and Code-Generation

NVIDIA • Town of Texas (WI)

On-site
USD 152,000 - 242,000
Stock options
Benefits
Paid time off
Senior AI Compiler Engineer, Algorithms and Code-Generation
Senior AI Compiler Engineer, Algorithms and Code-Generation

NVIDIA • Washington

On-site
USD 152,000 - 242,000
Equity
Benefits
Senior AI Compiler Engineer, Algorithms and Code-Generation
Senior AI Compiler Engineer, Algorithms and Code-Generation

NVIDIA • Austin (TX)

On-site
USD 152,000 - 242,000
Equity
Benefits
Senior AI Compiler Engineer, Algorithms and Code-Generation
Senior AI Compiler Engineer, Algorithms and Code-Generation

NVIDIA • California (MO)

On-site
USD 152,000 - 242,000
Senior AI Compiler Engineer - Applied Research
Senior AI Compiler Engineer - Applied Research

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 152,000 - 242,000
Equity
Generous benefits package
Senior AI Compiler Engineer, MLIR
Senior AI Compiler Engineer, MLIR

NVIDIA • Austin (TX)

On-site
USD 152,000 - 242,000
Equity
Benefits