Engineering Manager, Deep Learning Inference — GPU AI

NVIDIA AI

Santa Clara (CA)

On-site

USD 224,000 - 431,000

Full time

6 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity compensation
Benefits

Job summary

NVIDIA is seeking a Manager of Deep Learning Inference Software to lead a world-class team delivering GPU-accelerated inference software. You will guide OSS frameworks strategy, oversee optimization of large-scale models, and drive cross-functional collaboration across compiler, libraries, and research teams.

Ideal candidates have 6+ years of software development, 3+ years in leadership, and deep experience with CUDA, Triton, and CUTLASS. Equity and benefits are included.

Qualifications

  • MS, PhD, or equivalent experience in Computer Science, Electrical/Computer Engineering, or a related field.
  • 6+ overall years of software development experience, including 3+ years in technical leadership or engineering management.
  • Strong background in C/C++ software design and development; proficiency in Python is a plus.
  • Hands-on experience with GPU programming (CUDA, Triton, CUTLASS) and performance optimization.
  • Proven record of deploying or optimizing deep learning models in production environments.
  • Experience leading teams using Agile or collaborative software development practices.

Responsibilities

  • Lead, mentor, and scale a high-performing engineering team focused on deep learning inference and GPU-accelerated software.
  • Guide the strategy, roadmap, and execution of NVIDIA's OSS inference frameworks engineering.
  • Partner with internal compiler, libraries, and research teams to deliver end-to-end optimized inference pipelines across NVIDIA accelerators.
  • Oversee performance tuning, profiling, and optimization of large-scale models for LLM, multimodal, and generative AI applications.
  • Guide engineers in adopting best practices for CUDA, Triton, CUTLASS, and multi-GPU communications (NIXL, NCCL, NVSHMEM).
  • Represent the team in roadmap and planning discussions, ensuring alignment with NVIDIA's broader AI and software strategies.
  • Foster a culture of technical excellence, open collaboration, and continuous innovation.

Skills

Leadership
Agile practices
Mentoring teams

Education

MS/PhD or equivalent experience in CS/EE

Tools

CUDA
Triton
CUTLASS

Job description

NVIDIA is seeking a Manager of Deep Learning Inference Software to lead a world-class team delivering GPU-accelerated inference software. You will guide OSS frameworks strategy, oversee optimization of large-scale models, and drive cross-functional collaboration across compiler, libraries, and research teams.

Ideal candidates have 6+ years of software development, 3+ years in leadership, and deep experience with CUDA, Triton, and CUTLASS. Equity and benefits are included.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Engineering Manager, Deep Learning Inference & GPU
Engineering Manager, Deep Learning Inference & GPU

NVIDIA • Georgia

On-site
USD 224,000 - 432,000
Equity
Benefits
Engineering Manager, GPU AI Inference & Open-Source
Engineering Manager, GPU AI Inference & Open-Source

Nvidia Corporation • Santa Clara (CA)

On-site
USD 272,000 - 431,250
Equity
Benefits package
Engineering Manager: Deep Learning Inference & GPU
Engineering Manager: Deep Learning Inference & GPU

Jobtailor • California (MO)

On-site
USD 180,000 - 260,000
Engineering Manager: GPU-Accelerated AI Inference
Engineering Manager: GPU-Accelerated AI Inference

NVIDIA • Illinois

On-site
USD 224,000 - 432,000
Equity
Benefits
Engineering Manager, AI Inference — GPU-Accelerated DL
Engineering Manager, AI Inference — GPU-Accelerated DL

NVIDIA • California (MO)

On-site
USD 272,000 - 432,000
Engineering Manager – GPU Deep Learning Inference & Frameworks
Engineering Manager – GPU Deep Learning Inference & Frameworks

Socket.dev • Santa Clara (UT)

On-site
USD 224,000 - 357,000
Equity
Benefits package
Engineering Manager - GPU AI Inference & Frameworks
Engineering Manager - GPU AI Inference & Frameworks

NVIDIA • Town of Texas (WI)

On-site
USD 184,000 - 357,000
Equity
Benefits
Engineering Manager - Deep Learning Inference on GPUs
Engineering Manager - Deep Learning Inference on GPUs

NVIDIA • Massachusetts

On-site
USD 224,000 - 431,000
Equity compensation
Comprehensive benefits
Engineering Manager, GPU DL Inference & OSS Frameworks
Engineering Manager, GPU DL Inference & OSS Frameworks

NVIDIA • Santa Clara (CA)

Hybrid
USD 224,000 - 431,250
Head of GPU-Accelerated AI Inference
Head of GPU-Accelerated AI Inference

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 224,000 - 431,000
Equity
Benefits