Senior AI Performance & Efficiency Engineer - GPU Clusters

NVIDIA

Santa Clara (CA)

On-site

USD 152,000 - 241,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Comprehensive benefits package

Job summary

NVIDIA is seeking a Senior AI/ML Performance and Efficiency Engineer to improve AI efficiency on GPU Clusters. This role involves collaborating with AI researchers and automating the identification of inefficiencies in compute infrastructure.

The ideal candidate has over 5 years of experience in large-scale infrastructure, strong programming skills, and a dedication to staying current with AI technologies. Competitive salaries and comprehensive benefits are offered.

Qualifications

  • Minimum 5+ years of experience designing large scale compute infrastructure.
  • Strong understanding of modern ML techniques.
  • Experience with debugging large-scale distributed training.

Responsibilities

  • Collaborate with AI/ML researchers to enhance ML model efficiency.
  • Apply ML techniques to identify efficiency bottlenecks.
  • Monitor fleet-wide utilization patterns and identify improvements.

Skills

Collaboration with researchers
Efficiency in ML models
Debugging and optimization
Programming in Python
Experience with cloud computing

Education

BS in Computer Science or related area

Tools

NSight Systems
NSight Compute
AWS
GCP
Azure

Job description

NVIDIA is seeking a Senior AI/ML Performance and Efficiency Engineer to improve AI efficiency on GPU Clusters. This role involves collaborating with AI researchers and automating the identification of inefficiencies in compute infrastructure.

The ideal candidate has over 5 years of experience in large-scale infrastructure, strong programming skills, and a dedication to staying current with AI technologies. Competitive salaries and comprehensive benefits are offered.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Performance and Efficiency Engineer
Senior AI Performance and Efficiency Engineer

Nvidia Corporation • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Competitive salaries
Comprehensive benefits package
Equity eligibility
Senior AI Infrastructure Engineer — GPU Clusters
Senior AI Infrastructure Engineer — GPU Clusters

Nvidia Corporation • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior Full-Stack Engineer, AI Infra for GPU Clusters
Senior Full-Stack Engineer, AI Infra for GPU Clusters

NVIDIA • California (MO)

On-site
USD 184,000 - 357,000
Equity
Benefits
Lead AI Infrastructure Architect for Large-Scale GPU Clusters
Lead AI Infrastructure Architect for Large-Scale GPU Clusters

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 357,000
Equity and benefits
Senior AI Systems Tools Engineer - GPU Clusters
Senior AI Systems Tools Engineer - GPU Clusters

NVIDIA Gruppe • Santa Clara (CA)

Hybrid
USD 184,000 - 357,000
Senior AI Infrastructure Architect – GPU Clusters
Senior AI Infrastructure Architect – GPU Clusters

NVIDIA • California (MO)

On-site
USD 184,000 - 357,000
Equity
Benefits
Senior AI Infrastructure Engineer — Scalable GPU Clusters
Senior AI Infrastructure Engineer — Scalable GPU Clusters

NVIDIA AI • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Senior AI Infra Engineer — Scalable GPU Clusters
Senior AI Infra Engineer — Scalable GPU Clusters

2100 NVIDIA USA • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior AI GPU Cluster Performance Engineer
Senior AI GPU Cluster Performance Engineer

Socket.dev • Austin (TX)

Hybrid
USD 120,000 - 160,000
AMD benefits at a glance
Senior AI GPU Infra Engineer — Performance & Scale
Senior AI GPU Infra Engineer — Performance & Scale

Summit Group Solutions, LLC • United States

On-site
USD 150,000 - 350,000