Product Manager - BioNeMo Inference

NVIDIA AI

Santa Clara (CA)

On-site

USD 148,000 - 224,000

Full time

2 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA is seeking a technical Product Manager for BioNeMo Inference to define and drive deployment, operation, and scaling of biomolecular AI inference workloads across cloud and on-prem environments.

You will collaborate with engineering, research, and platform teams to create production-ready experiences, build APIs/SDKs, and set metrics for adoption and performance while balancing open-source models, NVIDIA NIMs, and enterprise deployment strategy.

Qualifications

  • 5+ years of industry experience, ideally including 3+ years in software engineering or a deeply technical role and 2+ years in product management.
  • Master's degree (or equivalent experience) in Electrical, Mechanical, Materials, or other related fields.
  • Demonstrated experience defining and shipping technical products, platforms, APIs, SDKs, developer tools, or cloud services.
  • Strong understanding of AI/ML inference systems and deployment tradeoffs, including model serving, GPU infrastructure, batching, throughput, latency, scaling, and cost-performance optimization.
  • Familiarity with containers and cloud-native infrastructure, including Docker, Kubernetes, Helm, CI/CD, and public-cloud deployment concepts.

Responsibilities

  • Define product vision, strategy, and roadmap for BioNeMo inference products, including NIM-based deployment, performance optimization, scalability, and developer onboarding.
  • Work closely with engineering, research, solution architects, cloud, and product teams to translate model capabilities into production-ready inference experiences.
  • Define requirements for inference optimization: batching, throughput, latency, GPU utilization, multi-GPU and multi-node scaling, caching, scheduling, observability, and reliability.
  • Partner with platform teams to ensure BioNeMo inference products deploy cleanly across cloud and enterprise environments, including Kubernetes-based environments.
  • Develop the developer experience across APIs, SDKs, containers, Helm charts, reference architectures, documentation, and examples.
  • Engage directly with early customers and partners to understand workflows, validate product direction, and turn feedback into prioritized requirements.
  • Establish product metrics for adoption, usability, performance, and operational quality; use data and customer insight to improve the product.
  • Drive cross-functional product launches with marketing, developer relations, sales, and solution architects.
  • Help shape the product boundary between open-source biomolecular models, NVIDIA NIMs, and enterprise-ready deployment and support!

Skills

AI/ML inference
Cloud-native infra
Docker/Kubernetes
Product management
Developer experience

Education

Master's degree in Electrical, Mechanical, Materials, or related field

Tools

NVIDIA NIM
Triton
TensorRT-LLM
vLLM
Ray
KServe
Kubeflow
Docker
Kubernetes
CI/CD
Helm

Job description

Job Requisition ID JR2020823

Job Category Marketing

Time Type Full time

NVIDIA is advancing the frontier of AI for biology with BioNeMo, bringing accelerated computing and generative AI to biomolecular research and drug discovery. We are seeking a technical Product Manager to lead BioNeMo Inference. You will define how developers, researchers, and enterprise platform teams deploy, operate, and scale biomolecular AI inference workloads. This role sits at the intersection of AI infrastructure, developer experience, and product execution: translating the needs of model developers and end users into simple, reliable inference products built on NVIDIA NIM and accelerated computing. A biology or healthcare background is not required.

We are looking for a strong technical PM who understands the fundamentals of AI inference serving and model deployment and is eager to apply them to a new and high-impact domain.

What You'll Be Doing
  • Define product vision, strategy, and roadmap for BioNeMo inference products, including NIM-based deployment, performance optimization, scalability, and developer onboarding.
  • Work closely with engineering, research, solution architects, cloud, and product teams to translate model capabilities into production-ready inference experiences.
  • Define requirements for inference optimization: batching, throughput, latency, GPU utilization, multi-GPU and multi-node scaling, caching, scheduling, observability, and reliability.
  • Partner with platform teams to ensure BioNeMo inference products deploy cleanly across cloud and enterprise environments, including Kubernetes-based environments.
  • Develop the developer experience across APIs, SDKs, containers, Helm charts, reference architectures, documentation, and examples.
  • Engage directly with early customers and partners to understand workflows, validate product direction, and turn feedback into prioritized requirements.
  • Establish product metrics for adoption, usability, performance, and operational quality; use data and customer insight to improve the product.
  • Drive cross-functional product launches with marketing, developer relations, sales, and solution architects.
  • Help shape the product boundary between open-source biomolecular models, NVIDIA NIMs, and enterprise-ready deployment and support!
What We Need To See
  • 5+ years of industry experience, ideally including 3+ years in software engineering or a deeply technical role and 2+ years in product management.
  • Master's degree (or equivalent experience) in Electrical, Mechanical, Materials, or other related fields.
  • Demonstrated experience defining and shipping technical products, platforms, APIs, SDKs, developer tools, or cloud services.
  • Strong understanding of AI/ML inference systems and deployment tradeoffs, including model serving, GPU infrastructure, batching, throughput, latency, scaling, and cost-performance optimization.
  • Familiarity with containers and cloud-native infrastructure, including Docker, Kubernetes, Helm, CI/CD, and public-cloud deployment concepts.
  • Ability to work fluently with engineers on system architecture, performance bottlenecks, operational requirements, and roadmap tradeoffs.
  • Strong customer discovery, product requirement definition, prioritization, and roadmap-management skills.
  • Excellent written and verbal communication skills for both technical and non-technical audiences.
  • A collaborative, self-starting approach and the ability to influence across highly matrixed teams.
Ways To Stand Out From The Crowd
  • Experience launching and scaling AI inference, model-serving, MLOps, or GPU-accelerated infrastructure products.
  • Hands-on expertise with modern AI inference and orchestration platforms such as NVIDIA NIM, Triton, TensorRT-LLM, vLLM, Ray, KServe, or Kubeflow.
  • Proven success optimizing production inference workloads, including performance, scalability, observability, and multi-tenant serving.
  • Experience building developer platforms or open-source products with strong ecosystem adoption and integrations.
  • Ability to partner with research teams to productize frontier AI models, including familiarity with biomolecular and scientific foundation models.

NVIDIA is widely considered to be one of the technology world's most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 148,000 USD - 224,250 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 8, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Product Manager - BioNeMo Inference
Product Manager - BioNeMo Inference

NVIDIA Corporation • Santa Clara (CA), Northern (KY)

Hybrid
USD 148,000 - 224,000
Equity
Benefits
Product Marketing Manager, NVIDIA NIM
Product Marketing Manager, NVIDIA NIM

NVIDIA • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Product Marketing Manager, NVIDIA NIM
Product Marketing Manager, NVIDIA NIM

NVIDIA AI • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Senior Product Manager – AI Inference Performance
Senior Product Manager – AI Inference Performance

NVIDIA • Santa Clara (CA)

On-site
USD 208,000 - 328,000
Product Manager – BioNeMo Inference & AI Deployment
Product Manager – BioNeMo Inference & AI Deployment

NVIDIA AI • Santa Clara (CA)

On-site
USD 148,000 - 224,000
Equity
Benefits
Product Manager - Open Models
Product Manager - Open Models

NVIDIA • Santa Clara (CA)

Hybrid
USD 168,000 - 328,000
Equity
Benefits
Solutions Architect - Drug Discovery
Solutions Architect - Drug Discovery

Nvidia Corporation • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Benefits
Solutions Architect - Drug Discovery
Solutions Architect - Drug Discovery

NVIDIA • California (MO)

On-site
USD 152,000 - 288,000
Solutions Architect - Drug Discovery
Solutions Architect - Drug Discovery

NVIDIA • Massachusetts

On-site
USD 152,000 - 288,000
Equity
Benefits
Solutions Architect - Drug Discovery
Solutions Architect - Drug Discovery

NVIDIA AI • California (MO)

On-site
USD 152,000 - 288,000
Equity
Benefits