N’envoyez pas un CV générique — générez un CV et une lettre de motivation adaptés à ce poste précis.
Nebius is advancing AI infrastructure with Token Factory, a high-performance GPU cloud platform. The role focuses on building and optimizing low-level inference kernels and runtime components for massive-scale AI workloads.
You will work closely with ML and backend teams to improve end-to-end execution, profile system and hardware performance, and integrate support for new GPU architectures to push efficiency and reliability across Nebius cloud services.
Nebius is advancing AI infrastructure with Token Factory, a high-performance GPU cloud platform. The role focuses on building and optimizing low-level inference kernels and runtime components for massive-scale AI workloads.
You will work closely with ML and backend teams to improve end-to-end execution, profile system and hardware performance, and integrate support for new GPU architectures to push efficiency and reliability across Nebius cloud services.