Inference Engineer — Low-Latency AI Pipelines

H Company

Paris

Hybride

EUR 60 000 - 90 000

Plein temps

14 jours+

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Avantages offerts par ce poste

Competitive salary
Opportunities for professional growth
Collaborative multicultural team

Résumé du poste

H Company is seeking a skilled AI Engineer to develop scalable inference pipelines in Paris. Candidates should possess a MS or PhD in Computer Science or Machine Learning, along with proficiency in Python, Rust, or C/C++. The role emphasizes collaboration with research teams and optimizing GPU performance for low-latency tasks. A competitive salary and opportunities for professional growth are offered. The position operates in a hybrid model requiring presence in the office three days a week.

Qualifications

  • Proficient in one of the programming languages: Python, Rust or C/C++.
  • Experience in GPU programming such as CUDA and Open AI Triton.
  • Experience in model compression and quantization techniques.

Responsabilités

  • Develop scalable, low-latency and cost effective inference pipelines.
  • Optimize model performance: memory usage, throughput, and latency.
  • Collaborate with H research teams on model architectures.

Connaissances

Programming in Python, Rust, or C/C++
GPU programming (CUDA, Open AI Triton)
Model compression and quantization techniques
Strong communication and presentation skills
Collaborative mindset

Formation

MS or PhD in Computer Science, Machine Learning

Outils

Pytorch
ONNX Runtime
CUDA

Description du poste

H Company is seeking a skilled AI Engineer to develop scalable inference pipelines in Paris. Candidates should possess a MS or PhD in Computer Science or Machine Learning, along with proficiency in Python, Rust, or C/C++. The role emphasizes collaboration with research teams and optimizing GPU performance for low-latency tasks. A competitive salary and opportunities for professional growth are offered. The position operates in a hybrid model requiring presence in the office three days a week.
Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

Senior AI Inference Systems Engineer (GPU & HPC)
Senior AI Inference Systems Engineer (GPU & HPC)

NVIDIA • France

Sur place
EUR 90 000 - 150 000
AI Engineer - Self-Hosted LLMs & High-Performance Inference
AI Engineer - Self-Hosted LLMs & High-Performance Inference

In Tandem • Paris

Sur place
EUR 60 000 - 80 000
Research Engineer, Model Inference & Serving - Paris
Research Engineer, Model Inference & Serving - Paris

H Company • Paris

Sur place
EUR 60 000 - 90 000
Competitive salary
Opportunities for professional growth
Collaborative multicultural team
Data Engineer: Scalable Pipelines for AI Safety
Data Engineer: Scalable Pipelines for AI Safety

European Tech Recruit • Paris

Hybride
EUR 60 000 - 80 000
Speech ML Research Engineer: Low-Latency, Scalable AI
Speech ML Research Engineer: Low-Latency, Scalable AI

Institute of Foundation Models • Paris

Sur place
EUR 45 000 - 75 000
Inference Performance Engineer
Inference Performance Engineer

adaption • Paris

Sur place
EUR 90 000 - 140 000
Flexible work
Adaption Passport
Lunch Stipend
+1
Paris-Based AI Engineer — Early-Stage, Impact & Growth
Paris-Based AI Engineer — Early-Stage, Impact & Growth

Amo • Paris

Sur place
EUR 70 000 - 90 000
100% health care coverage
8-9 weeks vacation per year
Sponsorship for visa process
+2
ML Inference Systems Engineer - Accelerator Performance
ML Inference Systems Engineer - Accelerator Performance

Arago • Paris

Sur place
EUR 120 000 - 180 000
Stock options
Healthcare coverage
Pension contributions
+2
Data Engineer: Build Scalable Pipelines for AI & ML
Data Engineer: Build Scalable Pipelines for AI & ML

The Resume Database • Paris

Sur place
EUR 50 000 - 70 000
AI Engineer: Scalable ML & AI Pipelines (Lille + Remote)
AI Engineer: Scalable ML & AI Pipelines (Lille + Remote)

OVENTI Consulting • Lille

Hybride
EUR 40 000 - 50 000
Télétravail 2 jours par semaine
Formation continue
Culture d'excellence