GPU Kernel Engineer for High-Performance AI Inference

Ensimag Alumni

Paris

Sur place

EUR 50 000 - 70 000

Plein temps

14 jours+

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Résumé du poste

Ensimag Alumni is seeking a talented engineer in Paris to work on GPU optimization for cutting-edge LLM inference at Kog. You'll engage in writing and profiling GPU kernels, optimizing computational paths, and contributing to a monokernel pipeline.

The ideal candidate has experience with CUDA or similar environments and a strong background from a top engineering school or relevant PhD. Join a dynamic team and push the boundaries of GPU engineering.

Qualifications

  • Experience in writing GPU kernels with performance constraints.
  • Understanding of hardware and GPU architecture.
  • Familiarity with inline PTX or CDNA ISA.

Responsabilités

  • Understand GPU internals and optimize computational sections.
  • Contribute to a monokernel pipeline for LLM inference.
  • Build profiling infrastructure for GPU programs.

Connaissances

GPU optimization
CUDA programming
PyTorch custom ops
Kernel optimization

Formation

Top engineering school or PhD with GPU work

Description du poste

Ensimag Alumni is seeking a talented engineer in Paris to work on GPU optimization for cutting-edge LLM inference at Kog. You'll engage in writing and profiling GPU kernels, optimizing computational paths, and contributing to a monokernel pipeline.

The ideal candidate has experience with CUDA or similar environments and a strong background from a top engineering school or relevant PhD. Join a dynamic team and push the boundaries of GPU engineering.

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

GPU Engineer
GPU Engineer

Ensimag Alumni • Paris

Sur place
EUR 50 000 - 70 000
Senior AI Inference Systems Engineer (GPU & HPC)
Senior AI Inference Systems Engineer (GPU & HPC)

NVIDIA • France

Sur place
EUR 90 000 - 150 000
ML Inference Systems Engineer — High-Performance Accelerator
ML Inference Systems Engineer — High-Performance Accelerator

Arago Inc. • Paris

Sur place
EUR 110 000 - 170 000
Stock options
Health insurance
Pension contributions
+1
ML Inference Systems Engineer - Accelerator Performance
ML Inference Systems Engineer - Accelerator Performance

Arago • Paris

Sur place
EUR 120 000 - 180 000
Stock options
Healthcare coverage
Pension contributions
+2
HPC & AI Computational Scientist - GPU Performance Engineer
HPC & AI Computational Scientist - GPU Performance Engineer

AMD • France

Sur place
EUR 90 000 - 130 000
GPU HPC & AI Scientist — Paris Center of Excellence
GPU HPC & AI Scientist — Paris Center of Excellence

AMD • Paris

Sur place
EUR 70 000 - 100 000
Senior Kernel Engineer: Custom Linux for Confidential AI
Senior Kernel Engineer: Custom Linux for Confidential AI

Chikara HR • Paris

Sur place
EUR 85 000 - 100 000
Own kernel equity
Publishable work
Leadership track in systems
+3
Research Engineer
Research Engineer

Kog • Paris

Hybride
EUR 60 000 - 90 000
Direct access to AMD and NVIDIA GPUs
Creative and decision-influencing environment
Remote-friendly working model
Senior Software Engineer, AI Inference Systems
Senior Software Engineer, AI Inference Systems

NVIDIA • France

Sur place
EUR 90 000 - 150 000
Ingénieur IA Temps Réel - GPU Embarqué (C++/CUDA)
Ingénieur IA Temps Réel - GPU Embarqué (C++/CUDA)

Safran.AI • Massy

Sur place
EUR 44 000 - 53 000
Participation et intéressement
Mutuelle familiale à 55 %
Jusqu’à 3 jours de télétravail par semaine