Senior Product Manager - AI Platform Inference

NVIDIA Corporation

Santa Clara (CA)

On-site

USD 168,000 - 328,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA Corporation in Santa Clara, CA is seeking a Senior Product Manager for AI Platform Inference to lead the development of tools, SDKs, and libraries enabling developers to deploy AI inference on NVIDIA GPUs.

The role emphasizes building products, defining roadmaps, and collaborating with developers to optimize model deployment performance, with strong emphasis on GenAI concepts and GPU-aware software delivery.

Qualifications

  • Experience with Inference deployment and optimization software.
  • Knowledge of GenAI or ML concepts, especially regarding performance optimization.
  • BS or MS in Computer Science, Computer Engineering, or equivalent experience, with 6+ years in technical product management.
  • Strong communication and interpersonal skills.

Responsibilities

  • Create products to help developers build better Inference deployments.
  • Develop product strategy, roadmaps, and go-to-market plans.
  • Collaborate with internal and external developers to build product roadmaps for model optimization software.
  • Work with leadership to align with and drive company strategy.

Skills

PM experience
GenAI knowledge
Communication
Software deployment

Education

BS or MS in CS/CE

Tools

vLLM
SGLang
FlashInfer
TensorRT-LLM
Triton
Dynamo
TorchAO

Job description

## Senior Product Manager - AI Platform InferenceApplylocations: US, CA, Santa Claratime type: Full timeposted on: Posted Yesterdayjob requisition id: JR2022724Inference is the fastest growing and most competitive area in Generative AI today. It is where AI models impact our daily life, and where ever bit of accuracy and performance matters for quality, safety, and cost. Inference is also constantly evolving, with new acceleration algorithms, usecases, and deployment techniques. As a Senior Product Manager for AI Platform Inference you will be responsible for building the tools, SDKs, and libraries which enables developers' Inference deployments to thrive on NVIDIA GPUs.As NVIDIA Product Managers, our goal is to enable developers to be successful on the NVIDIA Platform, and push the boundaries of what is possible with their AI deployments! For Inference, we are the champions inside NVIDIA for AI developers looking to accelerate their deployments on GPUs. We work directly with developers inside and outside of the company to identify key improvements, create roadmaps, and stay alert on the inference landscape. We also work with NVIDIA leaders to define clear product strategy, and marketing team teams to build go-to-market plans. The Product Management organization at NVIDIA is a small, strong, and impactful group. We focus on enabling deep learning across all GPU use cases and providing great solutions for developers. We are seeking a rare blend of product skills, technical depth, and passion to make NVIDIA great for developers. Does that sounds familiar? If so, we would love to hear from you!**What you'll be doing:*** Create products to help developers build better Inference deployments* Develop product strategy, roadmaps, and go-to-market plans* Collaborate with internal and external developers to build product-based roadmaps for model optimization software* Work with leadership to align with and drive company strategy**What we need to see:*** Experience with Inference deployment and optimization software (ex. vLLM, SGLang, FlashInfer, TensorRT-LLM, Triton, Dynamo, TorchAO, etc.)* Demonstrable knowledge of GenAI or machine learning concepts, particularly around performance optimization, and software development and delivery* BS or MS degree in Computer Science, Computer Engineering, or similar experience (or equivalent experience)* 6+ years of technical product management, or similar, experience at a technology company* Strong communication and interpersonal skills**Ways to stand out from the crowd:*** Experience leading optimization products for Inference* Working on Open Source & Github-first developer products with deep customer interactions* Knowledge of GPU architecture, HW/SW co-design, and performance profilingYour base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 168,000 USD - 258,750 USD for Level 4, and 208,000 USD - 327,750 USD for Level 5.You will also be eligible for equity and benefits.Applications for this job will be accepted at least until August 9, 2026.This posting is for an existing vacancy.NVIDIA uses AI tools in its recruiting processes.NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Product Manager - AI Platform Inference
Senior Product Manager - AI Platform Inference

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 168,000 - 328,000
Equity
Benefits
Senior Product Manager - AI Platform Inference
Senior Product Manager - AI Platform Inference

NVIDIA • Santa Clara (CA)

On-site
USD 208,000 - 328,000
Equity
Benefits
Senior Deep Learning Software Engineer, Inference
Senior Deep Learning Software Engineer, Inference

NVIDIA Corporation • Northern (KY)

On-site
USD 152,000 - 288,000
Senior Technical Product Marketing Manager - AI Infrastructure
Senior Technical Product Marketing Manager - AI Infrastructure

NVIDIA Corporation • Santa Clara (CA)

Hybrid
USD 152,000 - 230,000
Equity
Benefits
Senior DL Algorithms Engineer - Inference Performance
Senior DL Algorithms Engineer - Inference Performance

NVIDIA Corporation • Santa Clara (CA)

Hybrid
USD 184,000 - 288,000
Product Manager - Open Models
Product Manager - Open Models

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 168,000 - 328,000
Product Manager - Open Models
Product Manager - Open Models

NVIDIA Corporation • United States

Hybrid
USD 168,000 - 328,000
Equity
Benefits
Hybrid work
Senior AI Workflow Engineer
Senior AI Workflow Engineer

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

On-site
USD 184,000 - 357,000
Equity compensation
Health insurance
Relocation support
Engineering Manager, Deep Learning Inference
Engineering Manager, Deep Learning Inference

NVIDIA • Washington

On-site
USD 184,000 - 357,000
Equity
Benefits package
Senior Product Manager – AI Inference Performance
Senior Product Manager – AI Inference Performance

NVIDIA • California (MO)

On-site
USD 208,000 - 328,000