Remote AI Research Engineer: Kernel & Inference Optimization

Lever, Inc.

Netherlands

Remote

EUR 120,000 - 170,000

Full time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Remote-first team
International collaboration
Cutting-edge AI research projects
Flexible work environment

Job summary

Lever, Inc. in the Netherlands is seeking an AI Research Engineer (Kernel & Inference Optimization) to advance model-serving architectures across hardware environments.

You will tackle latency, throughput, and memory challenges on mobile and edge devices, developing GPU kernels and applying techniques like pruning and quantization.

This remote-first role combines research with systems engineering, collaborating with international teams to push efficient AI at scale.

Qualifications

  • PhD in NLP/ML with strong AI research track record.
  • Proven expertise in Metal Shading Language (MSL) and custom GPU kernels.
  • Experience with latency, throughput and memory optimization on mobile/edge hardware.
  • Experience with diffusion models and Vision Transformers.

Responsibilities

  • Design and deploy advanced model-serving architectures for high throughput and low latency.
  • Develop inference pipelines across mobile and edge environments.
  • Establish performance targets and run controlled benchmarks.
  • Develop custom GPU kernels and optimize memory usage.
  • Collaborate with cross-functional teams in a remote, highly technical setting.

Skills

Latency optimization
Inference benchmarking
Model serving architectures
Diffusion models
Vision Transformers
Distributed inference
English communication

Education

PhD in NLP/ML (relevant)

Tools

MSL (Metal Shading Language)
GPU kernel programming
Pruning/Quantization

Job description

Lever, Inc. in the Netherlands is seeking an AI Research Engineer (Kernel & Inference Optimization) to advance model-serving architectures across hardware environments.

You will tackle latency, throughput, and memory challenges on mobile and edge devices, developing GPU kernels and applying techniques like pruning and quantization.

This remote-first role combines research with systems engineering, collaborating with international teams to push efficient AI at scale.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Research Engineer (Kernel & Inference Optimization)
AI Research Engineer (Kernel & Inference Optimization)

Lever, Inc. • Netherlands

Remote
EUR 120,000 - 170,000
Remote-first team
International collaboration
Cutting-edge AI research projects
+1
Senior ML Engineer, LLM Inference Optimization
Senior ML Engineer, LLM Inference Optimization

Lever, Inc. • Netherlands

On-site
EUR 120,000 - 180,000
Competitive compensation
Career growth opportunities
Flexibility and ownership over work
+2
Edge AI Platform SRE: Scalable Cloud Infra (Remote Europe)
Edge AI Platform SRE: Scalable Cloud Infra (Remote Europe)

Qualcomm • Netherlands

Remote
EUR 90,000 - 120,000
Senior Machine Learning Engineer, LLM Inference Optimization
Senior Machine Learning Engineer, LLM Inference Optimization

Jobgether SRL • Netherlands

On-site
EUR 120,000 - 180,000
Competitive compensation
Career growth
Ownership over technical work
+1
Senior Machine Learning Engineer, LLM Inference Optimization
Senior Machine Learning Engineer, LLM Inference Optimization

Lever, Inc. • Netherlands

On-site
EUR 120,000 - 180,000
Competitive compensation
Career growth opportunities
Flexibility and ownership over work
+2
Senior Field Applications Engineer — AI Edge (Remote)
Senior Field Applications Engineer — AI Edge (Remote)

Innatera • Rijswijk

Hybrid
EUR 70,000 - 110,000
Flexible hours
Work-from-home policy
Inclusive culture
Autonomous AI Infra Engineer - GPU Fleet (Remote/UK)
Autonomous AI Infra Engineer - GPU Fleet (Remote/UK)

Together AI • Amsterdam

Hybrid
EUR 90,000 - 140,000
Senior AI Engineer I — Remote, Scalable AI Solutions
Senior AI Engineer I — Remote, Scalable AI Solutions

Vacaturebank • Netherlands

On-site
EUR 90,000 - 150,000
Remote work flexibility
Competitive salary
Comprehensive benefits
+1
Senior ML Engineer: LLM/VLM Inference Optimizer
Senior ML Engineer: LLM/VLM Inference Optimizer

Jobgether SRL • Netherlands

On-site
EUR 120,000 - 180,000
Competitive compensation
Career growth
Ownership over technical work
+1
Senior DL Researcher, On-Device AI & Model Efficiency
Senior DL Researcher, On-Device AI & Model Efficiency

Qualcomm • Amsterdam

On-site
EUR 120,000 - 170,000