Senior TensorRT Edge-LLM Inference Engineer

NVIDIA

California (MO)

On-site

USD 152,000 - 241,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA is seeking a software engineer to join its TensorRT Edge-LLM team, focusing on developing a state-of-the-art inference framework for large language models. The role requires a deep understanding of transformer models, modern C++ programming skills, and experience with LLM frameworks. The position offers competitive salary ranges between $152,000 and $287,500, equity, and benefits. NVIDIA values diversity and provides an equal-opportunity workplace.

Qualifications

  • 4+ years of relevant software development experience.
  • Familiarity with popular LLM frameworks and libraries.
  • A track record of strong software design and execution.

Responsibilities

  • Develop a state-of-the-art inference framework in modern C++.
  • Design and implement optimizations for transformer-based models.
  • Benchmark and optimize inference performance across environments.

Skills

Deep understanding of transformer models
Modern C++ programming (C++11/14/17)
Inference optimization techniques
Collaboration across fields

Education

BS, MS, PhD in Computer Science or related field

Tools

TensorRT
CUDA

Job description

NVIDIA is seeking a software engineer to join its TensorRT Edge-LLM team, focusing on developing a state-of-the-art inference framework for large language models. The role requires a deep understanding of transformer models, modern C++ programming skills, and experience with LLM frameworks. The position offers competitive salary ranges between $152,000 and $287,500, equity, and benefits. NVIDIA values diversity and provides an equal-opportunity workplace.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer – TensorRT Edge-LLM
Senior Software Engineer – TensorRT Edge-LLM

NVIDIA • California (MO)

On-site
USD 152,000 - 242,000
Equity
Benefits
Senior Edge AI Engineer: LLM Inference (TensorRT)
Senior Edge AI Engineer: LLM Inference (TensorRT)

NVIDIA Corporation • Santa Clara (CA)

Hybrid
USD 152,000 - 288,000
Equity
Benefits
Hybrid work model
Senior ML Inference Engineer – TensorRT & LLMs
Senior ML Inference Engineer – TensorRT & LLMs

NVIDIA • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior LLM Inference Architect — Edge, Data Center, Remote
Senior LLM Inference Architect — Edge, Data Center, Remote

Cerence AI • United States

Hybrid
USD 185,000 - 280,000
Annual bonus opportunity
Insurance coverage (medical, dental, vision, life, and disability)
Paid time off
Senior Software Engineer – TensorRT Edge-LLM
Senior Software Engineer – TensorRT Edge-LLM

NVIDIA Corporation • Santa Clara (CA)

Hybrid
USD 152,000 - 288,000
Equity
Benefits
Hybrid work model
Senior LLM Inference & Algorithms Engineer Remote, Equity
Senior LLM Inference & Algorithms Engineer Remote, Equity

NVIDIA • United States

On-site
USD 272,000 - 432,000
Equity
Benefits
Senior GPU AI Platforms Engineer - Edge LLM Inference
Senior GPU AI Platforms Engineer - Edge LLM Inference

NVIDIA • Durham (NC)

On-site
USD 224,000 - 357,000
Equity
Benefits
Senior LLM Inference Algorithms Engineer — Equity Options
Senior LLM Inference Algorithms Engineer — Equity Options

NVIDIA • California (MO)

On-site
USD 272,000 - 432,000
Equity
Benefits
Senior Software Engineer, DL Inference (TensorRT)
Senior Software Engineer, DL Inference (TensorRT)

NVIDIA AI • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Benefits
Manager, Large Language Model Inference
Manager, Large Language Model Inference

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 184,000 - 287,500
Competitive salary
Equity options
Comprehensive benefits