Senior DL Algorithms Engineer - Inference Performance

NVIDIA

California (MO)

On-site

USD 152,000 - 241,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA in California seeks a dedicated AI Software Engineer to enable and optimize open models using its advanced inference software stack. Ideal candidates should have a PhD in CS, EE, or equivalent, with at least 3 years of experience in deep learning and neural networks. Applicants will work on performance profiling and GPU optimization, contributing to open-source projects. The salary range is between 152,000 USD and 241,500 USD for Level 3, and 184,000 USD to 287,500 USD for Level 4, alongside equity and benefits.

Qualifications

  • PhD in a related field.
  • 3+ years of experience in deep learning and neural networks.
  • Proficient with AI or HPC frameworks.

Responsibilities

  • Optimize open models on NVIDIA’s accelerated software stack.
  • Deliver production code to open‑source frameworks.
  • Analyze bottlenecks across the inference stack.

Skills

Deep learning
Neural networks
Performance profiling
GPU optimization
PyTorch
Computer architecture

Education

PhD in CS, EE, CSEE or equivalent

Tools

CUDA
OpenCL

Job description

What You Will Be Doing
  • Enable and optimize state‑of‑the‑art open models (like Nemotron and Cosmos) on NVIDIA’s accelerated inference software stack.
  • Contribute new features, fix bugs and deliver production code to open‑source frameworks such as TRT‑LLM, vLLM, SGLang, FlashInfer, etc.
  • Profile and analyze bottlenecks across the full inference stack to push the boundaries of inference performance.
  • Benchmark state‑of‑the‑art offerings and perform competitive analysis for NVIDIA’s software/hardware stack.
  • Co‑design with partner teams to develop the next generation of AI models and services.
What We Want To See
  • PhD in CS, EE, CSEE or equivalent experience.
  • 3+ years of experience.
  • Strong background in deep learning and neural networks, especially inference.
  • Experience with performance profiling, analysis and optimization, especially for GPU‑based applications.
  • Proficient in PyTorch or equivalent frameworks for AI, or HPC‑heavy application development.
  • Deep understanding of computer architecture and familiarity with the fundamentals of GPU architecture.
Ways To Stand Out From The Crowd
  • Proven experience with processor and system‑level performance optimization.
  • Deep understanding of modern LLM/Diffusion architectures.
  • Strong fundamentals in algorithms.
  • GPU programming experience (CUDA or OpenCL) is a strong plus.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD – 241,500 USD for Level 3, and 184,000 USD – 287,500 USD for Level 4.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until May 9, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering a diverse work environment and proud to be an equal‑opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior DL Algorithms Engineer - Inference Performance
Senior DL Algorithms Engineer - Inference Performance

NVIDIA • Santa Clara (CA)

On-site
USD 152,000 - 242,000
Equity options
Comprehensive benefits
Engineering Manager, Deep Learning Inference
Engineering Manager, Deep Learning Inference

NVIDIA • Santa Clara (CA)

Hybrid
USD 224,000 - 431,250
Engineering Manager, Deep Learning Inference
Engineering Manager, Deep Learning Inference

Nvidia Corporation • Santa Clara (CA)

On-site
USD 272,000 - 431,250
Equity
Benefits package
Engineering Manager, Deep Learning Inference
Engineering Manager, Deep Learning Inference

2100 NVIDIA USA • Santa Clara (CA)

Hybrid
USD 224,000 - 432,000
Inference Performance Engineer, Agent Driven Inference Optimization
Inference Performance Engineer, Agent Driven Inference Optimization

NVIDIA AI • Santa Clara (CA)

On-site
USD 140,000 - 230,000
Inference Performance Engineer, AI Inference Configuration Optimization
Inference Performance Engineer, AI Inference Configuration Optimization

Nvidia Corporation • Santa Clara (CA)

On-site
USD 152,000 - 242,000
Inference Performance Engineer, Agent Driven Inference Optimization
Inference Performance Engineer, Agent Driven Inference Optimization

NVIDIA • California (MO)

On-site
USD 124,000 - 196,000
Equity eligibility
Engineering Manager, Deep Learning Inference
Engineering Manager, Deep Learning Inference

NVIDIA • Massachusetts

On-site
USD 224,000 - 431,000
Equity compensation
Comprehensive benefits
Inference Performance Engineer, AI Inference Configuration Optimization
Inference Performance Engineer, AI Inference Configuration Optimization

NVIDIA • Santa Clara (CA)

Hybrid
USD 124,000 - 242,000
Equity
Benefits package
Engineering Manager, Deep Learning Inference
Engineering Manager, Deep Learning Inference

NVIDIA • California (MO)

On-site
USD 272,000 - 432,000