Engineering Manager: Deep Learning Inference & GPU

Jobtailor

California (MO)

On-site

USD 180,000 - 260,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NVIDIA is seeking a senior engineering leader to guide an elite team focused on deep learning inference and GPU-accelerated software. You will shape strategy, drive execution of inference frameworks for Client AI, and collaborate with compiler, libraries, and research teams to optimize pipelines across accelerators.

You will mentor engineers, champion CUDA/Triton/CUTLASS adoption, and push performance optimization for large-scale models. Expect a roles-focused, impact-driven environment.

Qualifications

  • MS/PhD or equivalent in CS/EE or related field.
  • 6+ years of software development experience.
  • 3+ years in technical leadership or engineering management.
  • Strong background in C/C++ software design and development.
  • Proficiency in Python is a plus.
  • Hands-on GPU programming with CUDA/Triton/CUTLASS.
  • Experience with performance optimization in production.
  • Leading teams with Agile or collaborative practices.
  • Open-source contributions to DL or inference frameworks advantageous.

Responsibilities

  • Lead, mentor, and scale a high-performing engineering team focused on deep learning inference and GPU-accelerated software.
  • Drive strategy, roadmap, and execution of inference frameworks engineering, focusing on Client AI.
  • Partner with compiler, libraries, and research teams to deliver optimized inference pipelines across NVIDIA accelerators.
  • Oversee performance tuning, profiling, and optimization of large-scale models for LLM and multimodal AI.
  • Guide engineers in adopting CUDA, Triton, and multi-GPU communication technologies.
  • Represent the team in roadmap and planning discussions and foster technical excellence.

Skills

Technical Leadership
Deep Learning Model Optimization
CUDA Programming
Performance Tuning
Agile Software Development
Team Leadership

Education

MS/PhD in CS/EE or related field

Tools

CUDA
Triton
CUTLASS
NIXL
NCCL
NVSHMEM

Job description

NVIDIA is seeking a senior engineering leader to guide an elite team focused on deep learning inference and GPU-accelerated software. You will shape strategy, drive execution of inference frameworks for Client AI, and collaborate with compiler, libraries, and research teams to optimize pipelines across accelerators.

You will mentor engineers, champion CUDA/Triton/CUTLASS adoption, and push performance optimization for large-scale models. Expect a roles-focused, impact-driven environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Engineering Manager, Deep Learning Inference & GPU
Engineering Manager, Deep Learning Inference & GPU

NVIDIA • Georgia

On-site
USD 224,000 - 432,000
Equity
Benefits
Engineering Manager, Deep Learning Inference — GPU AI
Engineering Manager, Deep Learning Inference — GPU AI

NVIDIA AI • Santa Clara (CA)

On-site
USD 224,000 - 431,000
Equity compensation
Benefits
Engineering Manager - Deep Learning Inference on GPUs
Engineering Manager - Deep Learning Inference on GPUs

NVIDIA • Massachusetts

On-site
USD 224,000 - 431,000
Equity compensation
Comprehensive benefits
Head of GPU-Accelerated AI Inference
Head of GPU-Accelerated AI Inference

Nvidia Corporation in • Santa Clara (CA)

On-site
USD 224,000 - 431,000
Equity
Benefits
Engineering Manager, AI Inference — GPU-Accelerated DL
Engineering Manager, AI Inference — GPU-Accelerated DL

NVIDIA • California (MO)

On-site
USD 272,000 - 432,000
Head of GPU-Accelerated AI Inference
Head of GPU-Accelerated AI Inference

NVIDIA • Washington

On-site
USD 224,000 - 432,000
Equity
Benefits package
Engineering Manager: GPU-Accelerated AI Inference
Engineering Manager: GPU-Accelerated AI Inference

NVIDIA • Illinois

On-site
USD 224,000 - 432,000
Equity
Benefits
Engineering Manager - GPU AI Inference & Frameworks
Engineering Manager - GPU AI Inference & Frameworks

NVIDIA • Town of Texas (WI)

On-site
USD 184,000 - 357,000
Equity
Benefits
Engineering Manager – GPU Deep Learning Inference & Frameworks
Engineering Manager – GPU Deep Learning Inference & Frameworks

Socket.dev • Santa Clara (UT)

On-site
USD 224,000 - 357,000
Equity
Benefits package
Engineering Manager, GPU AI Inference & Open-Source
Engineering Manager, GPU AI Inference & Open-Source

Nvidia Corporation • Santa Clara (CA)

On-site
USD 272,000 - 431,250
Equity
Benefits package