Principal LLM Inference Algorithms Engineer

NEPSE Trading

Northern (KY)

Hybrid

USD 272,000 - 431,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA seeks outstanding engineers to advance LLM inference, optimizing agentic workloads and building scalable, datacenter-scale systems. You will contribute to performance-enhancing algorithms and protocols that push the boundaries of what's possible with large language models.

Required: advanced degrees and 15+ years in deep learning, strong Python/C++ skills, and deep knowledge of GPU/parallel computing. Equity and benefits accompany the role. Applications accepted until July 26, 2026.

Qualifications

  • BS, MS or PhD in CS/EE/CE or related field.
  • 15+ years of deep learning and systems design experience.
  • Proficiency in Python and C++ programming.
  • Strong understanding of computer architecture and GPU computing.

Responsibilities

  • Research and develop generative AI, agents, and inference systems into the NVIDIA LLM software stack.
  • Analyze workloads and optimize agentic LLM workloads to reduce latency and increase throughput.
  • Design scalable systems to accelerate agentic workflows for datacenter-scale use cases.
  • Collaborate with diverse teams and external partners to formalize requirements.

Skills

Python
C++
Deep learning systems
Performance modeling

Education

BS/MS/PhD in Computer Science, Electrical Engineering, Computer Engineering, or related field

Tools

CUDA
OpenCL
GPU architecture

Job description

NVIDIA seeks outstanding engineers to advance LLM inference, optimizing agentic workloads and building scalable, datacenter-scale systems. You will contribute to performance-enhancing algorithms and protocols that push the boundaries of what's possible with large language models.

Required: advanced degrees and 15+ years in deep learning, strong Python/C++ skills, and deep knowledge of GPU/parallel computing. Equity and benefits accompany the role. Applications accepted until July 26, 2026.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Principal Deep Learning Algorithm Engineer
Principal Deep Learning Algorithm Engineer

NEPSE Trading • Northern (KY)

Hybrid
USD 272,000 - 431,000
Equity
Benefits
Senior DL Inference Engineer - GPU/LLM Performance & Equity
Senior DL Inference Engineer - GPU/LLM Performance & Equity

NVIDIA • Washington

On-site
USD 184,000 - 288,000
Equity
Benefits
Senior Deep Learning Software Engineer, Inference
Senior Deep Learning Software Engineer, Inference

NVIDIA • United States

Remote
USD 184,000 - 357,000
Senior Deep Learning Software Engineer, Inference
Senior Deep Learning Software Engineer, Inference

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Equity
Benefits
LLM Inference Performance Engineer
LLM Inference Performance Engineer

Intel • Santa Clara (CA)

Hybrid
USD 171,000 - 315,000
Senior Deep Learning Software Engineer, Inference
Senior Deep Learning Software Engineer, Inference

NVIDIA Gruppe • California (MO)

On-site
USD 152,000 - 288,000
Senior Deep Learning Software Engineer, Inference
Senior Deep Learning Software Engineer, Inference

NVIDIA AI • Washington

On-site
USD 140,000 - 230,000
Equity
Comprehensive benefits package
Senior Deep Learning Software Engineer, Inference
Senior Deep Learning Software Engineer, Inference

NVIDIA • Washington

On-site
USD 184,000 - 288,000
Equity
Benefits
Senior Inference Engineer, GPU Kernel Optimization
Senior Inference Engineer, GPU Kernel Optimization

Nvidia Corporation • Santa Clara (CA)

On-site
USD 184,000 - 288,000
Equity
Comprehensive benefits
Senior Deep Learning Inference Engineer - Equity Eligible
Senior Deep Learning Inference Engineer - Equity Eligible

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 224,000 - 357,000
Equity
Benefits