Senior DL Inference Engineer — GPU-Optimized AI at Scale

NVIDIA Corporation

Netherlands

Remote

EUR 120,000 - 180,000

Full time

6 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

NVIDIA seeks a Senior Software Engineer specializing in Deep Learning Inference for our growing team. You will design, build, and optimize the GPU-accelerated software that powers today’s AI applications, with focus on SGLang and vLLM.

You will collaborate across frameworks and libraries, deliver performance improvements for state-of-the-art models on datacenter GPUs and edge devices, and contribute to NVIDIA’s inference ecosystem including CUDA kernels and multi-GPU communication tools.

Qualifications

  • Masters or PhD or equivalent experience in Computer Engineering, Computer Science, EECS or AI.
  • 5+ years of relevant software development experience.
  • Excellent C/C++ programming and software design skills.
  • SW Agile skills are helpful; Python experience is a plus.
  • Prior experience with training, deploying or optimizing the inference of DL models in production is a plus.
  • Background with performance modeling, profiling, debugging and code optimization or architectural knowledge of CPU and GPU is a plus.

Responsibilities

  • Performance optimization and tuning of DL models in domains such as LLM, Multimodal and Generative AI.
  • Scale performance of DL models across architectures and NVIDIA accelerators.
  • Contribute features and code to NVIDIA’s inference libraries, vLLM and SGLang, FlashInfer and LLM software solutions.
  • Work with cross-collaborative teams across frameworks, NVIDIA libraries and inference optimization innovative solutions.

Skills

C/C++ programming
Python
Software design
Agile

Education

Master's/PhD in Computer Science/Engineering/AI

Tools

CUDA
CUTLASS
NCCL
vLLM
SGLang
OAI Triton

Job description

NVIDIA seeks a Senior Software Engineer specializing in Deep Learning Inference for our growing team. You will design, build, and optimize the GPU-accelerated software that powers today’s AI applications, with focus on SGLang and vLLM.

You will collaborate across frameworks and libraries, deliver performance improvements for state-of-the-art models on datacenter GPUs and edge devices, and contribute to NVIDIA’s inference ecosystem including CUDA kernels and multi-GPU communication tools.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior DL Inference Engineer - GPU-Accelerated AI
Senior DL Inference Engineer - GPU-Accelerated AI

NVIDIA • Amsterdam

On-site
EUR 120,000 - 180,000
Senior Deep Learning Software Engineer, Inference
Senior Deep Learning Software Engineer, Inference

NVIDIA • Amsterdam

On-site
EUR 120,000 - 180,000
Senior Deep Learning Software Engineer, Inference
Senior Deep Learning Software Engineer, Inference

NVIDIA Corporation • Netherlands

On-site
EUR 120,000 - 180,000
Senior Deep Learning Engineer, Accuracy Evaluation
Senior Deep Learning Engineer, Accuracy Evaluation

NVIDIA • Netherlands

On-site
EUR 120,000 - 180,000
Senior Machine Learning Engineer, LLM Inference Optimization
Senior Machine Learning Engineer, LLM Inference Optimization

Jobgether SRL • Netherlands

On-site
EUR 120,000 - 180,000
Competitive compensation
Career growth
Ownership over technical work
+1
Senior GPU ML Engineer - Inference Kernel Optimizer
Senior GPU ML Engineer - Inference Kernel Optimizer

Slashhash • Amsterdam

Hybrid
EUR 120,000 - 150,000
Senior ML Platform Engineer — Scalable Inference
Senior ML Platform Engineer — Scalable Inference

Atlassian • Amsterdam

Hybrid
EUR 120,000 - 180,000
Health resources
Volunteer days
Community engagement
Senior ML Engineer — Scale LLM Inference on GPU Cloud
Senior ML Engineer — Scale LLM Inference on GPU Cloud

United States Digital Space LLC • Amsterdam

Hybrid
EUR 100,000 - 180,000
Competitive compensation
Career growth and learning
Flexibility and ownership
+3
Senior ML Engineer — High-Performance AI Inference
Senior ML Engineer — High-Performance AI Inference

Nebius • Amsterdam

On-site
EUR 70,000 - 90,000
Competitive compensation
Career growth opportunities
Collaborative culture
+1
Senior Machine Learning Engineer, LLM Inference Optimization
Senior Machine Learning Engineer, LLM Inference Optimization

Lever, Inc. • Netherlands

On-site
EUR 120,000 - 180,000
Competitive compensation
Career growth opportunities
Flexibility and ownership over work
+2