Senior AI Systems Engineer - Inference & GPU Kernels

NVIDIA

Westford (MA)

On-site

USD 184,000 - 288,000

Full time

8 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA seeks outstanding AI systems engineers to advance the inference software stack for high-impact workloads. You will develop libraries, code generators, and GPU kernel tech for NVIDIA hardware, building abstractions for LLM serving and just-in-time compilers across frameworks and runtimes.

Required: Masters in CS/EE (PhD preferred), 6+ years in ML/DL systems, strong Python/C++, and CUDA-based GPU kernel work. Eligible for equity and benefits as part of the package.

Qualifications

  • Masters degree in Computer Science, Electrical Engineering, or related field (or equivalent experience); PhD preferred.
  • 6+ years (academic/ industry) experience with ML/DL systems development preferable.
  • Strong experience in developing or using deep learning frameworks (e.g. PyTorch, JAX, TensorFlow, ONNX, etc) and ideally inference engines and runtimes such as vLLM, SGLang, and MLC.
  • Strong Python and C/C++ programming skills
  • Strong experience in GPU kernel development and performance optimizations (especially using CUDA C/C++, cuTile, Triton, or similar) with hands-on experience with Matrix Multiplication.

Responsibilities

  • Innovating and developing new AI systems technologies for efficient inference
  • Designing, implementing, and optimizing kernels for high impact AI workloads
  • Designing and implementing extensible abstractions for LLM serving engines
  • Building efficient just-in-time domain specific compilers and runtimes
  • Collaborating closely with other engineers at NVIDIA across deep learning frameworks, libraries, kernels, and GPU arch teams
  • Contributing to open source communities like FlashInfer, vLLM, and SGLang

Skills

Python programming
C/C++ programming
GPU kernel development
CUDA
Performance optimization
ML/DL systems development

Education

Masters degree in Computer Science or Electrical Engineering
PhD preferred

Tools

PyTorch
JAX
TensorFlow
ONNX
vLLM
SGLang
Triton
MLIR

Job description

NVIDIA seeks outstanding AI systems engineers to advance the inference software stack for high-impact workloads. You will develop libraries, code generators, and GPU kernel tech for NVIDIA hardware, building abstractions for LLM serving and just-in-time compilers across frameworks and runtimes.

Required: Masters in CS/EE (PhD preferred), 6+ years in ML/DL systems, strong Python/C++, and CUDA-based GPU kernel work. Eligible for equity and benefits as part of the package.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Systems Engineer - Inference & GPU Kernels
Senior AI Systems Engineer - Inference & GPU Kernels

NVIDIA AI • California (MO)

On-site
USD 184,000 - 288,000
Senior AI Inference Systems Engineer – GPU Kernels
Senior AI Inference Systems Engineer – GPU Kernels

NVIDIA • Seattle (WA)

On-site
USD 184,000 - 288,000
Senior AI Inference & Kernel Engineer
Senior AI Inference & Kernel Engineer

NVIDIA • Durham (NC)

On-site
USD 184,000 - 288,000
Equity
Benefits
Senior AI Inference Engineer: GPU Kernels & LLM Runtimes
Senior AI Inference Engineer: GPU Kernels & LLM Runtimes

NVIDIA • Redmond (WA)

On-site
USD 184,000 - 288,000
Equity
Benefits
Senior AI Inference Systems Engineer | GPU Kernels & Runtime
Senior AI Inference Systems Engineer | GPU Kernels & Runtime

NVIDIA • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Senior AI Inference Systems Engineer (GPU & HPC)
Senior AI Inference Systems Engineer (GPU & HPC)

NVIDIA • Santa Clara (CA)

Hybrid
USD 184,000 - 357,000
Equity
Benefits
Senior AI Inference Engineer — GPU-Accelerated DL Systems
Senior AI Inference Engineer — GPU-Accelerated DL Systems

NVIDIA Corporation • California (MO)

On-site
USD 152,000 - 287,500
Equity
Benefits
Engineering Manager, Deep Learning Inference — GPU AI
Engineering Manager, Deep Learning Inference — GPU AI

NVIDIA AI • Santa Clara (CA)

On-site
USD 224,000 - 431,000
Equity compensation
Benefits
Engineering Manager, GPU AI Inference & Open-Source
Engineering Manager, GPU AI Inference & Open-Source

Nvidia Corporation • Santa Clara (CA)

On-site
USD 272,000 - 431,250
Equity
Benefits package
Senior ML Inference Engineer - GPUs & TensorRT
Senior ML Inference Engineer - GPUs & TensorRT

NVIDIA AI • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits