Senior AI Training Performance Architect

NVIDIA Corporation

Santa Clara (CA)

On-site

USD 210,000 - 340,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA Corporation in Santa Clara, CA is seeking a Senior AI Training Performance Architect to push the limits of AI training workloads. You will analyze, profile, and optimize across the hardware/software stack to extract peak performance for large-scale DL systems.

Join a team delivering production-quality software, contribute to MLPerf submissions, and influence roadmap decisions with deep expertise in deep learning, GPU architecture, and performance modeling.

Qualifications

  • PhD in CS, EE or CSEE (or equivalent experience) with 5+ years of relevant experience; or MS with 8+ years of experience.
  • Strong background in deep learning and neural networks, particularly in training.
  • Solid understanding of computer architecture and familiarity with GPU architecture fundamentals.
  • Proven background in analyzing and tuning application performance.
  • Proven experience with processor and system-level performance modeling.
  • Proficiency in programming with C++, Python, and CUDA.

Responsibilities

  • Understand, analyze, profile, and optimize AI training workloads on state-of- the hardware and software platforms.
  • Identifying performance bottlenecks of AI training on GPUs, prioritizing and then solving problems across the key AI training workloads.
  • Implement production-quality software across multiple layers of NVIDIA's deep learning platform stack, from drivers to DL frameworks.
  • Build and support NVIDIA submissions for MLPerf Training benchmarks.
  • Implement key DL training workloads in NVIDIA's proprietary processor and system simulators to enable future architecture studies.
  • Develop tools to automate workload analysis, optimization, and other critical workflows.

Skills

Deep learning
Neural networks
GPU architecture
Performance tuning
C++
Python
CUDA

Education

PhD in CS/EE/CSEE (or equivalent)
MS in CS/EE/CSEE (or equivalent)

Job description

We are now looking for a Senior AI Training Performance Architect NVIDIA is seeking a senior engineer who is obsessed with performance analysis and optimization to help us squeeze every last clock cycle out of AI training, the workload driving the design and construction of the largest and most powerful compute systems in the world. If you are willing to work across all layers of the hardware/software stack - from GPU architecture to the application code - to achieve peak performance, we want to hear from you. This role offers the opportunity to directly impact the hardware and software roadmap in a fast-growing technology company that leads the AI revolution. Join us and help design and build the world's most powerful compute systems!

What you will be doing
  • Understand, analyze, profile, and optimize AI training workloads on state-of- the hardware and software platforms.
  • Identifying performance bottlenecks of AI training on GPUs, prioritizing and then solving problems across the key AI training workloads.
  • Implement production-quality software across multiple layers of NVIDIA's deep learning platform stack, from drivers to DL frameworks.
  • Build and support NVIDIA submissions for MLPerf Training benchmarks.
  • Implement key DL training workloads in NVIDIA's proprietary processor and system simulators to enable future architecture studies.
  • Develop tools to automate workload analysis, optimization, and other critical workflows.
What we want to see
  • PhD in CS, EE or CSEE (or equivalent experience) with 5+ years of relevant experience; or MS with 8+ years of experience.
  • Strong background in deep learning and neural networks, particularly in training.
  • Solid understanding of computer architecture and familiarity with GPU architecture fundamentals.
  • Proven background in analyzing and tuning application performance.
  • Proven experience with processor and system-level performance modeling.
  • Proficiency in programming with C++, Python, and CUDA.
Benefits and Compensation
  • Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.
  • You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 28, 2026.

This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.

Equal Opportunity Employer

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer.

As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

NVIDIA pioneered accelerated computing. Today, our AI infrastructure powers global intelligence, transforming every industry. Learn more about NVIDIA.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Training Performance Architect
Senior AI Training Performance Architect

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 356,500
Equity
Benefits
Senior High-Performance LLM Training Engineer
Senior High-Performance LLM Training Engineer

NVIDIA Corporation • Santa Clara (CA)

Hybrid
USD 184,000 - 357,000
Equity
Benefits package
Hybrid work environment
Senior Deep Learning Performance Architect
Senior Deep Learning Performance Architect

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Senior AI and Quant DevTech Engineer
Senior AI and Quant DevTech Engineer

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior AI Performance and Efficiency Engineer
Senior AI Performance and Efficiency Engineer

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Senior Data Center Performance Engineer - Benchmarking and Optimization
Senior Data Center Performance Engineer - Benchmarking and Optimization

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Senior Solutions Architect, AI Performance Engineering
Senior Solutions Architect, AI Performance Engineering

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 356,500
Equity
Benefits package
Senior Software Engineer, CUDA Deep Learning Systems
Senior Software Engineer, CUDA Deep Learning Systems

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Equity
Benefits
Principal High-Performance LLM Training Engineer
Principal High-Performance LLM Training Engineer

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 272,000 - 431,000
Equity
Benefits
Senior Solutions Architect, AI Performance Engineering
Senior Solutions Architect, AI Performance Engineering

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 357,000