Senior GPU AI Platform Engineer — Edge Inference (Equity)

NVIDIA AI

Seattle (WA)

On-site

USD 224,000 - 431,250

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Benefits

Job summary

NVIDIA is seeking a seasoned engineer to advance LLM inference on edge AI hardware. You’ll evaluate frameworks, map model architectures to GPU, and optimize across multi-node deployments, focusing on performance and stability.

You’ll own validation workflows, develop actionable recipes, and collaborate with the community and partners to resolve hardware-specific inference issues. Equity and benefits are offered.

Qualifications

  • BS, MS, or PhD in Computer Science, Computer Engineering, Electrical Engineering, or equivalent experience.
  • 12+ years of software engineering with depth in GPU computing, ML systems, or high-performance inference.
  • Strong Python or C++ programming, software design, and software engineering skills.

Responsibilities

  • Track and evaluate innovations in open-source LLM inference frameworks and identify performance features.
  • Map new model architectures and inference algorithms to NVIDIA GPU architecture and optimize.
  • Characterize multi-node inference behavior and topology-aware all-reduce strategies on edge clusters.
  • Publish performance analyses mapping hardware limits to observed throughput, latency, and utilization.
  • Own the model validation workflow for new model releases and develop inference recipes.
  • Develop and maintain developer-facing inference recipes, automate staleness detection, and update CI results.
  • Engage with community and partners on model bring-up questions; act as technical point of contact for hardware-specific inference issues.

Skills

Python
C++
GPU computing
ML systems
Performance optimization

Education

BS/MS/PhD in CS/CE/EE or related field

Tools

CUDA
Triton

Job description

NVIDIA is seeking a seasoned engineer to advance LLM inference on edge AI hardware. You’ll evaluate frameworks, map model architectures to GPU, and optimize across multi-node deployments, focusing on performance and stability.

You’ll own validation workflows, develop actionable recipes, and collaborate with the community and partners to resolve hardware-specific inference issues. Equity and benefits are offered.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior GPU ML Inference Engineer — Edge AI Platforms
Senior GPU ML Inference Engineer — Edge AI Platforms

NVIDIA • Westford (MA)

On-site
USD 224,000 - 432,000
Senior GPU AI Platforms Engineer - Edge LLM Inference
Senior GPU AI Platforms Engineer - Edge LLM Inference

NVIDIA • Durham (NC)

On-site
USD 224,000 - 357,000
Equity
Benefits
GPU Inference Performance Engineer — Equity & Optimization
GPU Inference Performance Engineer — Equity & Optimization

Nvidia Corporation • Santa Clara (CA)

On-site
USD 152,000 - 242,000
Senior AI Systems Engineer – Edge GPU & Local Inference
Senior AI Systems Engineer – Edge GPU & Local Inference

NVIDIA • Santa Clara (CA)

On-site
USD 152,000 - 288,000
Equity
Benefits
AI Inference Platform Lead – Equity Eligible
AI Inference Platform Lead – Equity Eligible

NVIDIA • Washington

On-site
USD 184,000 - 357,000
Equity
Benefits package
Senior AI Inference Systems Engineer (GPU & HPC)
Senior AI Inference Systems Engineer (GPU & HPC)

NVIDIA • Santa Clara (CA)

Hybrid
USD 184,000 - 357,000
Equity
Benefits
Inference Performance Engineer: AI GPU Optimization&Equity
Inference Performance Engineer: AI GPU Optimization&Equity

NVIDIA • Santa Clara (CA)

Hybrid
USD 124,000 - 242,000
Equity
Benefits package
GPU Inference Engineer — Deep Learning
GPU Inference Engineer — Deep Learning

2100 NVIDIA USA • California (MO)

On-site
USD 124,000 - 242,000
Equity
Benefits
Senior Inference Performance Engineer — Equity & Hybrid
Senior Inference Performance Engineer — Equity & Hybrid

NVIDIA Gruppe • Santa Clara (CA)

Hybrid
USD 124,000 - 242,000
Senior AI Compiler Engineer - MLIR, GPU Inference, Equity
Senior AI Compiler Engineer - MLIR, GPU Inference, Equity

NVIDIA • Town of Texas (WI)

On-site
USD 152,000 - 288,000
Equity
Benefits package