Recevez plus de réponses des employeurs
Envoyez un CV adapté au poste en quelques minutes.
Arago Inc. in Paris is seeking a senior ML inference engineer to optimize execution on a custom accelerator. You will work across kernels, model execution, and multi-device distribution to maximize performance and efficiency.
You will contribute to the software stack, with emphasis on CUDA/Triton-based kernels, operator fusion, and advanced inference-serving techniques; strong C++/Python skills and English proficiency are required.
Arago Inc. in Paris is seeking a senior ML inference engineer to optimize execution on a custom accelerator. You will work across kernels, model execution, and multi-device distribution to maximize performance and efficiency.
You will contribute to the software stack, with emphasis on CUDA/Triton-based kernels, operator fusion, and advanced inference-serving techniques; strong C++/Python skills and English proficiency are required.