ML Inference Performance Intern — Benchmarking & Tools

Axelera AI

Eindhoven

Hybrid

EUR 42,000 - 64,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Pension plan
Employee insurances
Company shares

Job summary

Axelera AI is seeking a curious, rigorous engineer to join our team and dig into the performance of ML inference systems. You’ll work across the full inference stack—model export, compiler toolchains, and runtime execution on silicon—to build a clear, evidence-based picture of how different platforms perform and why.

Your work will go beyond running benchmarks: you’ll develop a repeatable evaluation methodology, investigate bottlenecks at hardware and software levels, and build tooling that

Qualifications

  • Currently enrolled in final years of Bachelor's or Master's programme or Masters thesis project.
  • Proficiency in Python and C/C++ development.
  • Experience with end-to-end computer vision pipelines.
  • Familiarity with benchmarking concepts (performance, latency).
  • Experience with inference tools, APIs, or SDKs (e.g., TensorRT).
  • Understanding of deep learning model concepts (quantization, ONNX, PyTorch).
  • Linux, Bash scripting, and Docker proficiency.
  • Hands-on experience with embedded hosts.
  • Strong written and verbal English communication.
  • Good organisational skills.

Responsibilities

  • Benchmarking & tooling: develop benchmarking tools for throughput, latency, power, and accuracy across platforms; maintain dashboards for visualisations.
  • Platform evaluation: evaluate AI accelerators and vendor SDKs and track model support across platforms.
  • Pipeline analysis: characterize full inference pipelines and ensure consistent configurations across platforms.
  • Lab & infrastructure: set up lab hosts across hardware platforms and onboard new evaluation hardware.
  • Reporting: synthesize findings into clear reports informing engineering and roadmap decisions.

Skills

Python
C/C++
CV pipelines
Benchmarking
Inference SDKs
DL concepts
Agentic AI
Git
LLM benchmarking
Linux
Bash scripting
Embedded hardware
English communication
Organisational skills

Education

Bachelor's or Master's in Computer Engineering/Computer Science/EE

Tools

TensorRT
GStreamer
Docker
Git

Job description

Axelera AI is seeking a curious, rigorous engineer to join our team and dig into the performance of ML inference systems. You’ll work across the full inference stack—model export, compiler toolchains, and runtime execution on silicon—to build a clear, evidence-based picture of how different platforms perform and why.

Your work will go beyond running benchmarks: you’ll develop a repeatable evaluation methodology, investigate bottlenecks at hardware and software levels, and build tooling that

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Inference Performance Engineer Intern
ML Inference Performance Engineer Intern

Axelera • Eindhoven

Hybrid
EUR 42,000 - 62,000
Pension plan
Employee share options
Insurance package
Intern - ML Inference Performance Engineer
Intern - ML Inference Performance Engineer

Axelera • Eindhoven

Hybrid
EUR 42,000 - 62,000
Pension plan
Employee share options
Insurance package
Intern - ML Inference Performance Engineer
Intern - ML Inference Performance Engineer

Axelera AI • Eindhoven

Hybrid
EUR 42,000 - 64,000
Pension plan
Employee insurances
Company shares
Senior ML Platform Engineer — Scalable Inference
Senior ML Platform Engineer — Scalable Inference

Atlassian • Amsterdam

Hybrid
EUR 120,000 - 180,000
Health resources
Volunteer days
Community engagement
Senior ML Engineer — High-Performance AI Inference
Senior ML Engineer — High-Performance AI Inference

Nebius • Amsterdam

On-site
EUR 70,000 - 90,000
Competitive compensation
Career growth opportunities
Collaborative culture
+1
ML Training Systems Engineer - Performance Focus
ML Training Systems Engineer - Performance Focus

Internetwork Expert • Amsterdam

Hybrid
EUR 100,000 - 150,000
High base salary
Generous bonus structure
Cutting-edge hardware and software
ML Performance Engineer
ML Performance Engineer

Internetwork Expert • Amsterdam

Hybrid
EUR 100,000 - 150,000
High base salary
Generous bonus structure
Cutting-edge hardware and software
AI Systems Engineer: Agents & Inference (Remote)
AI Systems Engineer: Agents & Inference (Remote)

Axelera AI • Netherlands

Hybrid
EUR 90,000 - 140,000
Pension plan
Employee insurances
Stock options
+1
Senior GPU ML Engineer - Inference Kernel Optimizer
Senior GPU ML Engineer - Inference Kernel Optimizer

Slashhash • Amsterdam

Hybrid
EUR 120,000 - 150,000
Sr. Machine Learning Engineer | Up to €170K
Sr. Machine Learning Engineer | Up to €170K

Bluebird • Amsterdam

On-site
EUR 144,000 - 170,000