Senior Research Engineer - Enterprise Products

NVIDIA

Seattle (WA)

On-site

USD 192,000 - 356,500

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA in Seattle is seeking a Senior Research Engineer focused on Generative AI inference. You will design routing policies for LLM traffic to balance accuracy and latency, and build agentic benchmarks to calibrate models across NVIDIA’s serving stack.

Collaborate with research and engineering teams to contribute design docs, code reviews, and open-source contributions, while mentoring junior engineers and driving experimental validation.

Qualifications

  • Bachelor's or Master's degree in Computer Science or equivalent.
  • 8+ years of industry experience in Deep Learning frameworks (PyTorch or TensorFlow).
  • Experience designing or running LLM evaluations/benchmarks and drawing sound conclusions.
  • Understanding of ML, DNNs, NLP, or Speech Recognition.
  • Empirical research mindset: form hypotheses, calibrate, iterate.
  • Strong communication and teamwork; mentoring junior engineers is a plus.
  • Desire to grow and learn new things.
  • Strong CS fundamentals: algorithms, data structures, parallel and distributed computing.

Responsibilities

  • Design and evaluate routing policies for LLM traffic to optimize model usage.
  • Build and run agentic benchmarks to measure algorithm quality and calibration data.
  • Ship to an open-source repo with design docs, code reviews, and community contributions.
  • Collaborate with engineering teams to integrate software across NVIDIA's serving stack.

Skills

Deep learning frameworks experience
LLM benchmarking
Machine learning fundamentals
Distributed & parallel computing
Algorithms & data structures
CUDA / GPU knowledge
Mentoring experience

Education

Bachelor's or Master's in Computer Science

Tools

CUDA
PyTorch
TensorFlow

Job description

We are now looking for a Senior Research Engineer passionate about Generative AI inference. Are you excited to change the way people infuse AI into products and services? NVIDIA is at the forefront of generative AI models, from language to images. NVIDIA provides building blocks to democratize AI and make generative AI easy to develop, integrate, and deploy. Our team is dedicated to developing optimized inferencing technologies to support our growing generative AI needs. We contribute to all steps of the machine learning lifecycle: from conceptualization, to applied research, engineering for optimized inference, and deployment. Collaborate with research teams, engineers, and open-source community.

What You Will Be Doing
  • Design and evaluate routing policies for LLM traffic to best use mixture of model systems.
  • Build and run agentic benchmarks (e.g., Terminal-Bench) to measure algorithm quality, and turn results into calibration data and routing profiles.
  • Ship to an open-source repo: design docs, code review, docs, and community contributions.
  • Collaborating with engineering teams across all of NVIDIA to ensure our software integrates seamlessly up and down the NVIDIA accelerated serving stack.
What We Need To See
  • Bachelor's or Master's degree in Computer Science or equivalent experience.
  • 8+ years of industry experience in Deep Learning frameworks (PyTorch or TensorFlow).
  • Experience designing or running LLM evaluations/benchmarks — ideally agentic ones — and drawing statistically sound conclusions from them.
  • Understanding of modern techniques in Machine Learning, Deep Neural Networks, Natural Language Processing, or Speech Recognition.
  • Empirical research mindset: forming hypotheses about new algorithms, running calibrations, iterating on results.
  • Strong communication and interpersonal skills, along with the ability to work in a dynamic and distributed team. A history of mentoring junior engineers and interns is a huge plus.
  • A desire to constantly grow and learn new things.
  • Strong computer science fundamentals - algorithms and data structures, computational complexity, parallel and distributed computing, system software.
Ways To Stand Out From a Crowd
  • Experience architecting or developing large-scale distributed systems for deep learning.
  • Agentic benchmark creation and publications.
  • Knowledge of CPU and/or GPU architecture.
  • GPU programming (CUDA).

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 192,000 USD - 304,750 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.

You will also be eligible for equity and benefits.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Research Engineer - Enterprise Products
Senior Research Engineer - Enterprise Products

NVIDIA Gruppe • Washington

On-site
USD 192,000 - 356,500
Equity
Benefits
Senior Deep Learning Software Engineer, Inference
Senior Deep Learning Software Engineer, Inference

NVIDIA • California (MO)

On-site
USD 152,000 - 287,500
Equity
Benefits package
Competitive salary
Engineering Manager, Agentic GenAI Platform
Engineering Manager, Agentic GenAI Platform

NVIDIA • Santa Clara (CA)

On-site
USD 224,000 - 432,000
Equity
Benefits
Principal Deep Learning Algorithm Engineer
Principal Deep Learning Algorithm Engineer

NVIDIA • California (MO)

On-site
USD 272,000 - 432,000
Equity
Benefits
Engineering Manager, Deep Learning Inference
Engineering Manager, Deep Learning Inference

Nvidia Corporation • Santa Clara (CA)

On-site
USD 272,000 - 431,250
Equity
Benefits package
Engineering Manager, Deep Learning Inference
Engineering Manager, Deep Learning Inference

NVIDIA • Georgia

On-site
USD 224,000 - 432,000
Equity
Benefits
Engineering Manager, Deep Learning Inference
Engineering Manager, Deep Learning Inference

NVIDIA • Massachusetts

On-site
USD 224,000 - 431,000
Equity compensation
Comprehensive benefits
Engineering Manager, Deep Learning Inference
Engineering Manager, Deep Learning Inference

NVIDIA • Illinois

On-site
USD 224,000 - 432,000
Equity
Benefits
Senior Solutions Architect, Generative AI Research
Senior Solutions Architect, Generative AI Research

NVIDIA • Georgia

On-site
USD 184,000 - 288,000
Engineering Manager, Deep Learning Inference
Engineering Manager, Deep Learning Inference

NVIDIA • California (MO)

On-site
USD 272,000 - 432,000