ML Inference Engineer | Accelerate AI on Custom Accelerator

Arago

Paris

Sur place

EUR 90 000 - 130 000

Plein temps

Il y a 43 heures
Soyez parmi les premiers à postuler

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Avantages offerts par ce poste

Stock options
Healthcare coverage
Pension contributions
Professional development
25 days PTO

Résumé du poste

Arago is seeking an experienced engineer to optimize AI model execution on its custom accelerator. You will work across kernels, model execution, multi-device distribution, and inference serving, shaping the software stack around the hardware capabilities.

You will contribute to high-performance ML inference, kernel optimization, and distributed execution while collaborating with hardware, compiler, and runtime teams to drive performance improvements.

Qualifications

  • Strong experience in high-performance ML inference and accelerator programming.
  • Deep understanding of computer architecture, memory hierarchies, and parallelism.
  • Experience developing and optimizing custom kernels using low-level environments (CUDA, Triton, ROCm/HIP).
  • Experience with operator fusion, tiling, scheduling, data movement optimization, and graph execution.

Responsabilités

  • Analyze modern AI workloads to identify kernel-, runtime-, memory-, and system-level bottlenecks on Arago's accelerator.
  • Develop and optimize custom kernels and fused operators for maximum device utilization.
  • Design efficient mappings of models across multiple Arago devices and manage communication/synchronization.
  • Develop inference-serving techniques such as continuous batching and KV caches.
  • Build profiling, benchmarking, and performance-analysis infrastructure spanning kernels and full models.

Connaissances

ML inference
GPU programming
Computer architecture
CUDA/Triton/ROCm
Kernels optimization
Distributed execution
Inference serving
C++
Python
MLIR
English proficiency

Outils

CUDA
TensorRT
ROCm
MLIR

Description du poste

Arago is seeking an experienced engineer to optimize AI model execution on its custom accelerator. You will work across kernels, model execution, multi-device distribution, and inference serving, shaping the software stack around the hardware capabilities.

You will contribute to high-performance ML inference, kernel optimization, and distributed execution while collaborating with hardware, compiler, and runtime teams to drive performance improvements.

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

ML Inference Systems Engineer — High-Performance Accelerator
ML Inference Systems Engineer — High-Performance Accelerator

Arago Inc. • Paris

Sur place
EUR 110 000 - 170 000
Stock options
Health insurance
Pension contributions
+1
ML Inference Systems Engineer - Accelerator Performance
ML Inference Systems Engineer - Accelerator Performance

Arago • Paris

Sur place
EUR 120 000 - 180 000
Stock options
Healthcare coverage
Pension contributions
+2
ML Systems Engineer — Inference Acceleration
ML Systems Engineer — Inference Acceleration

Arago • Paris

Sur place
EUR 120 000 - 180 000
Stock options
Healthcare coverage
Pension contributions
+2
ML Systems Engineer — Inference Acceleration
ML Systems Engineer — Inference Acceleration

Arago Inc. • Paris

Sur place
EUR 110 000 - 170 000
Stock options
Health insurance
Pension contributions
+1
ML Systems Engineer — Inference Acceleration
ML Systems Engineer — Inference Acceleration

Arago • Paris

Sur place
EUR 90 000 - 130 000
Stock options
Healthcare coverage
Pension contributions
+2
AI Compiler Engineer: Optimize ML on Custom Accelerators
AI Compiler Engineer: Optimize ML on Custom Accelerators

IC Resources • Grenoble

Sur place
EUR 90 000 - 120 000
AI Compiler Engineer for Custom AI Accelerators
AI Compiler Engineer for Custom AI Accelerators

IC Resources • Auvergne-Rhône-Alpes

Hybride
EUR 65 000 - 90 000
AI Compiler Engineer for Next-Gen AI Accelerators
AI Compiler Engineer for Next-Gen AI Accelerators

Kalray • Montbonnot-Saint-Martin

Sur place
EUR 90 000 - 120 000
Competitive salary
RTT extra leave
Sustainable mobility incentives
+2
Senior AI/ML Scientist
Senior AI/ML Scientist

Arago • Paris

Sur place
EUR 80 000 - 110 000
Competitive cash compensation
Stock options
Relocation bonus
+4
Senior AI/ML Scientist
Senior AI/ML Scientist

Arago Inc. • Paris

Sur place
EUR 70 000 - 100 000
Competitive cash compensation
Stock option plan
Healthcare coverage
+2