Senior High-Performance LLM Training Engineer

NVIDIA Corporation

Santa Clara (CA)

Hybrid

USD 184,000 - 357,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Equity
Benefits package
Hybrid work environment

Job summary

NVIDIA Corporation seeks a Senior High-Performance LLM Training Engineer to optimize AI training workloads and accelerate performance across thousands of GPUs.

You will profile, analyze, and implement production-quality software across the deep learning stack, from drivers to frameworks, while contributing to MLPerf submissions and architecture studies. Proficiency in PyTorch, CUDA, and C++ is essential.

Qualifications

  • PhD in Computer Science, Electrical Engineering or Computer Engineering or MS with extensive experience.
  • Strong background in deep learning and neural networks, especially training.
  • Proven experience analyzing and tuning application performance; system-level modelling.
  • Programming skills in C++, Python, and CUDA.

Responsibilities

  • Understand, profile, and optimize AI training workloads on advanced hardware and software platforms.
  • Analyze training performance across GPUs and neural networks; prioritize and solve problems.
  • Implement production-quality software across NVIDIA's deep learning stack, from drivers to frameworks.
  • Build and support MLPerf Training benchmark submissions.
  • Develop workloads for future architecture studies in simulators.
  • Create tools to automate workload analysis and optimization workflows.

Skills

C++
Python
CUDA
Deep learning
Neural networks
Performance tuning

Education

PhD in Computer Science, Electrical Engineering or Computer Engineering
MS in a related field

Job description

Senior High-Performance LLM Training Engineer

NVIDIA is seeking experienced engineers specializing in performance analysis and optimization to improve the efficiency of LLM training workloads, which are shaping the world's most advanced computing systems.

This position focuses on optimizing NVIDIA’s high-performance LLM software stack in frameworks like PyTorch and JAX for high-performance training on thousands of GPUs, while also helping shape hardware roadmaps for the next generation of GPUs powering the AI revolution.

What you will be doing:
  • Understand, analyze, profile, and optimize AI training workloads on innovative hardware and software platforms.
  • Understand the big picture of training performance on GPUs, prioritizing and then solving problems across all state-of-the-art neural networks.
  • Implement production-quality software in multiple layers of NVIDIA's deep learning platform stack, from drivers to DL frameworks.
  • Build and support NVIDIA submissions to the MLPerf Training benchmark suite.
  • Implement key DL training workloads in NVIDIA's proprietary processor and system simulators to enable future architecture studies.
  • Build tools to automate workload analysis, workload optimization, and other critical workflows.
What we want to see:
  • PhD in Computer Science, Electrical Engineering or Computer Engineering and 5+ years; or MS (or equivalent experience) and 8+ years of meaningful work experience.
  • Strong background in deep learning and neural networks, in particular training.
  • A deep background in computer architecture and familiarity with the fundamentals of GPU architecture.
  • Proven experience analyzing and tuning application performance & processor and system-level performance modelling.
  • Programming skills in C++, Python, and CUDA.

GPU computing is the most productive and pervasive platform for deep learning and AI. It begins with the most advanced GPUs and the systems and software we build on top of them. We integrate and optimize every deep learning framework. We work with the major systems companies and every major cloud service provider to make GPUs available in data centers and in the cloud. We craft computers and software to bring AI to edge devices, such as self-driving cars and autonomous robots. AI has the potential to spur a wave of social progress unmatched since the industrial revolution.

Widely considered to be one of tech's most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. Additionally, this opportunity offers you the ability to collaborate with some of the most forward-thinking and hard-working people in the world, shaping the future of AI in a creative and autonomous work environment that encourages innovation.

If you're excited to work across the full hardware & software stack—from GPU architecture to application code—to achieve optimal performance, we want to hear from you!

Salary

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5. You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until April 12, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

Equal Opportunity Employment

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

NVIDIA pioneered accelerated computing. Today, our AI infrastructure powers global intelligence, transforming every industry. Learn more about NVIDIA.

#LI-Hybrid

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior High-Performance LLM Training Engineer
Senior High-Performance LLM Training Engineer

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 356,500
Equity
Benefits
Principal High-Performance LLM Training Engineer
Principal High-Performance LLM Training Engineer

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 272,000 - 431,000
Equity
Benefits
Senior AI Training Performance Architect
Senior AI Training Performance Architect

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 210,000 - 340,000
Equity
Benefits
Principal High-Performance LLM Training Engineer
Principal High-Performance LLM Training Engineer

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 272,000 - 431,250
Senior AI Training Performance Architect
Senior AI Training Performance Architect

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 356,500
Equity
Benefits
Principal High-Performance LLM Training Engineer
Principal High-Performance LLM Training Engineer

NVIDIA • Santa Clara (CA)

On-site
USD 272,000 - 431,250
Equity
Benefits
Senior High-Performance LLM Training Engineer
Senior High-Performance LLM Training Engineer

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 287,500
Equity
Comprehensive benefits
Senior Solutions Architect, GPU Performance and LLM - Cloud Service Providers
Senior Solutions Architect, GPU Performance and LLM - Cloud Service Providers

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Senior Software Engineer, DGX Cloud AI Infrastructure
Senior Software Engineer, DGX Cloud AI Infrastructure

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Senior Software Engineer, CUDA Deep Learning Systems
Senior Software Engineer, CUDA Deep Learning Systems

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Equity
Benefits