Senior ML Inference Engineer - TensorRT/CUDA

NVIDIA Corporation

Santa Clara (CA)

Hybrid

USD 152,000 - 288,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA Corporation is seeking a Senior Software Engineer for the TensorRT team to advance inference software for AI accelerators. You will design, develop, and optimize TensorRT and TensorRT-LLM, writing C++, Python, and CUDA.

Collaborate with DL experts and GPU architects to influence hardware and software design for fast, efficient inference. We value creativity, collaboration and a passion for learning in a fast-paced environment and offer hybrid work arrangement and a competitive

Qualifications

  • BS, MS, PhD or equivalent experience in Computer Science, Computer Engineering or a related field.
  • 4+ years of software development experience on a large codebase or project.
  • Strong proficiency in C++ (required), Rust or Python programming languages.
  • Experience in developing Deep Learning Frameworks, Compilers, or System Software.
  • Excellent problem-solving skills and passion to learn and work effectively in a fast-paced, collaborative environment.

Responsibilities

  • Design, develop and optimize NVIDIA TensorRT and TensorRT-LLM to supercharge inference applications for datacenter, workstations, and PCs.
  • Develop software in C++, Python, and CUDA for seamless and efficient deployment of state-of-the-art LLMs and Generative AI models.
  • Collaborate with deep learning experts and GPU architects throughout the company to influence Hardware and Software design for inference.

Skills

C++
Rust
Python
Problem solving
Communication

Education

BS/MS/PhD or equivalent

Tools

CUDA
TensorRT
PyTorch
OpenCL

Job description

NVIDIA Corporation is seeking a Senior Software Engineer for the TensorRT team to advance inference software for AI accelerators. You will design, develop, and optimize TensorRT and TensorRT-LLM, writing C++, Python, and CUDA.

Collaborate with DL experts and GPU architects to influence hardware and software design for fast, efficient inference. We value creativity, collaboration and a passion for learning in a fast-paced environment and offer hybrid work arrangement and a competitive

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior ML Inference Engineer - GPUs & TensorRT
Senior ML Inference Engineer - GPUs & TensorRT

NVIDIA AI • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior ML Inference Engineer – TensorRT & LLMs
Senior ML Inference Engineer – TensorRT & LLMs

NVIDIA • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior DL Inference Engineer (TensorRT) - Hybrid + Equity
Senior DL Inference Engineer (TensorRT) - Hybrid + Equity

NVIDIA Corporation • Santa Clara (CA)

Hybrid
USD 152,000 - 288,000
Equity
Benefits
Senior Software Engineer, DL Inference (TensorRT)
Senior Software Engineer, DL Inference (TensorRT)

NVIDIA AI • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Benefits
Senior DL Inference Engineer — GPU-Optimized LLMs (Remote)
Senior DL Inference Engineer — GPU-Optimized LLMs (Remote)

NVIDIA Corporation • Northern (KY)

Hybrid
USD 152,000 - 288,000
Senior Software Engineer, Machine Learning Inference
Senior Software Engineer, Machine Learning Inference

NVIDIA AI • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Senior AI Inference Engineer — GPU-Accelerated DL Systems
Senior AI Inference Engineer — GPU-Accelerated DL Systems

NVIDIA Corporation • California (MO)

On-site
USD 152,000 - 287,500
Equity
Benefits
AI-Native Systems Engineer, TensorRT — Hybrid & Equity
AI-Native Systems Engineer, TensorRT — Hybrid & Equity

NVIDIA Corporation • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
Hybrid work model
Senior AI Inference Optimization Engineer
Senior AI Inference Optimization Engineer

Nvidia Corporation in • Santa Clara (CA)

Hybrid
USD 124,000 - 196,000
Equity
Benefits package
Hybrid work model
Senior Software Engineer, Machine Learning Inference
Senior Software Engineer, Machine Learning Inference

NVIDIA • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits