Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Luma AI is hiring for a Performance Optimization role focused on making multimodal models faster and more scalable. You will work closely with research and engineering to optimize training and inference, write high-performance kernels, and push transformer models toward lower latency and higher throughput.
The role emphasizes deep CUDA/Triton expertise, PyTorch proficiency, and the ability to deploy efficient, production-grade optimizations across hardware platforms.
Luma AI is hiring for a Performance Optimization role focused on making multimodal models faster and more scalable. You will work closely with research and engineering to optimize training and inference, write high-performance kernels, and push transformer models toward lower latency and higher throughput.
The role emphasizes deep CUDA/Triton expertise, PyTorch proficiency, and the ability to deploy efficient, production-grade optimizations across hardware platforms.