Senior LLM Training Performance Engineer (Hybrid)

NVIDIA Corporation

Santa Clara (CA)

Hybrid

USD 184,000 - 357,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Equity
Benefits package
Hybrid work environment

Job summary

NVIDIA Corporation seeks a Senior High-Performance LLM Training Engineer to optimize AI training workloads and accelerate performance across thousands of GPUs.

You will profile, analyze, and implement production-quality software across the deep learning stack, from drivers to frameworks, while contributing to MLPerf submissions and architecture studies. Proficiency in PyTorch, CUDA, and C++ is essential.

Qualifications

  • PhD in Computer Science, Electrical Engineering or Computer Engineering or MS with extensive experience.
  • Strong background in deep learning and neural networks, especially training.
  • Proven experience analyzing and tuning application performance; system-level modelling.
  • Programming skills in C++, Python, and CUDA.

Responsibilities

  • Understand, profile, and optimize AI training workloads on advanced hardware and software platforms.
  • Analyze training performance across GPUs and neural networks; prioritize and solve problems.
  • Implement production-quality software across NVIDIA's deep learning stack, from drivers to frameworks.
  • Build and support MLPerf Training benchmark submissions.
  • Develop workloads for future architecture studies in simulators.
  • Create tools to automate workload analysis and optimization workflows.

Skills

C++
Python
CUDA
Deep learning
Neural networks
Performance tuning

Education

PhD in Computer Science, Electrical Engineering or Computer Engineering
MS in a related field

Job description

NVIDIA Corporation seeks a Senior High-Performance LLM Training Engineer to optimize AI training workloads and accelerate performance across thousands of GPUs.

You will profile, analyze, and implement production-quality software across the deep learning stack, from drivers to frameworks, while contributing to MLPerf submissions and architecture studies. Proficiency in PyTorch, CUDA, and C++ is essential.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior LLM Training Performance Architect
Senior LLM Training Performance Architect

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 272,000 - 431,000
Equity
Benefits
Senior LLM Training Performance Architect
Senior LLM Training Performance Architect

NVIDIA • Santa Clara (CA)

Hybrid
USD 184,000 - 356,500
Equity
Benefits
Senior High-Performance LLM Training Engineer
Senior High-Performance LLM Training Engineer

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 356,500
Equity
Benefits
Senior High-Performance LLM Training Engineer
Senior High-Performance LLM Training Engineer

NVIDIA Corporation • Santa Clara (CA)

Hybrid
USD 184,000 - 357,000
Equity
Benefits package
Hybrid work environment
Lead High-Performance LLM Training Architect (Equity)
Lead High-Performance LLM Training Architect (Equity)

NVIDIA • Santa Clara (CA)

On-site
USD 272,000 - 431,250
Equity
Benefits
Principal High-Performance LLM Training Engineer
Principal High-Performance LLM Training Engineer

NVIDIA • Santa Clara (CA)

On-site
USD 272,000 - 431,250
Equity
Benefits
Senior Performance Engineer: AI/LLM Benchmark Lead
Senior Performance Engineer: AI/LLM Benchmark Lead

NVIDIA Corporation • Santa Clara (CA)

Hybrid
USD 136,000 - 270,000
Equity compensation
Benefits
Principal High-Performance LLM Training Engineer
Principal High-Performance LLM Training Engineer

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 272,000 - 431,000
Equity
Benefits
Senior AI Training Performance Architect — Optimize at Scale
Senior AI Training Performance Architect — Optimize at Scale

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 210,000 - 340,000
Equity
Benefits
Senior AI Inference Performance Engineer (CUDA/LLM/VLM)
Senior AI Inference Performance Engineer (CUDA/LLM/VLM)

NVIDIA AI • Santa Clara (CA)

On-site
USD 180,000 - 300,000
Equity
Generous Benefits Package