Senior DL Inference & Optimization Engineer

NL02 NVIDIA Dutch B.V. Incorporated

Netherlands

On-site

EUR 120,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NVIDIA seeks a Senior Software Engineer specializing in Deep Learning Inference for our growing team. As a key contributor, you will help design, build, and optimize the GPU-accelerated software that powers today’s most sophisticated AI applications.

Our team is responsible for developing and maintaining high-performance deep learning frameworks, including SGLang and vLLM, which are at the forefront of efficient large-scale model serving and inference.

Qualifications

  • Master's degree or PhD in Computer Science, Computer Engineering, EECS, or AI.
  • 5+ years of relevant software development experience.
  • Excellent C/C++ programming and software design skills; Python experience is a plus.
  • Prior experience with training, deploying or optimizing the inference of DL models in production is a plus.
  • Background with performance modeling, profiling, debugging, and CPU/GPU optimization is a plus.

Responsibilities

  • Performance optimization, analysis, and tuning of DL models in domains like LLM, multimodal, and Generative AI.
  • Scale performance of DL models across architectures and NVIDIA accelerators.
  • Contribute features and code to NVIDIA’s inference libraries, vLLM and SGLang, FlashInfer and LLM software solutions.
  • Work with cross-collaborative teams across frameworks, NVIDIA libraries and inference optimization innovations.

Skills

C/C++ programming
Python programming
Software design
Agile

Education

Master's or PhD in Computer Science/Engineering/AI

Tools

SGLang
vLLM
CUTLASS
OAI Triton
NCCL
CUDA kernels

Job description

NVIDIA seeks a Senior Software Engineer specializing in Deep Learning Inference for our growing team. As a key contributor, you will help design, build, and optimize the GPU-accelerated software that powers today’s most sophisticated AI applications.

Our team is responsible for developing and maintaining high-performance deep learning frameworks, including SGLang and vLLM, which are at the forefront of efficient large-scale model serving and inference.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior DL Inference Engineer - GPU-Accelerated AI
Senior DL Inference Engineer - GPU-Accelerated AI

NVIDIA • Amsterdam

On-site
EUR 120,000 - 180,000
Senior Deep Learning Software Engineer, Inference
Senior Deep Learning Software Engineer, Inference

NVIDIA • Amsterdam

On-site
EUR 120,000 - 180,000
Senior Deep Learning Software Engineer, Inference, Senior Deep Learning Software Engineer, Inference
Senior Deep Learning Software Engineer, Inference, Senior Deep Learning Software Engineer, Inference

NVIDIA • Netherlands

On-site
EUR 75,000 - 100,000
Highly competitive salaries
Extensive benefits package
Diversity and inclusion initiatives
Senior DL Inference Engineer - Scale & Performance
Senior DL Inference Engineer - Scale & Performance

NVIDIA • Netherlands

On-site
EUR 75,000 - 100,000
Senior Deep Learning Engineer: GPU Inference & Production
Senior Deep Learning Engineer: GPU Inference & Production

NVIDIA • Netherlands

On-site
EUR 221,000 - 507,000
Senior AI Evaluation Engineer for Frontier Models
Senior AI Evaluation Engineer for Frontier Models

NVIDIA • Netherlands

On-site
EUR 120,000 - 180,000
Senior Deep Learning Engineer, Accuracy Evaluation
Senior Deep Learning Engineer, Accuracy Evaluation

NVIDIA • Netherlands

On-site
EUR 120,000 - 180,000
Senior ML Engineer — High-Performance AI Inference
Senior ML Engineer — High-Performance AI Inference

Nebius • Amsterdam

On-site
EUR 70,000 - 90,000
Competitive compensation
Career growth opportunities
Collaborative culture
+1
Senior SRE — AI Inference Platform & GPU Cloud
Senior SRE — AI Inference Platform & GPU Cloud

United States Digital Space LLC • Amsterdam

Hybrid
EUR 120,000 - 160,000
Competitive compensation
Career growth and learning
Flexibility and ownership
+3
Staff Software Engineer, AI Inference Infrastructure
Staff Software Engineer, AI Inference Infrastructure

Together AI • Amsterdam

On-site
EUR 90,000 - 130,000