Get more replies from employers
Send a job-specific resume in minutes.
MBZUAI in Sunnyvale, CA seeks a Distributed ML Engineer to optimize ML software stacks for training and inference on state-of-the-art hardware. You will prototype kernels, build benchmarks, and help scale large‑scale models while collaborating with researchers and engineers.
The role emphasizes parallel computing, system‑level coding and hands-on ML experience, with visa sponsorship available. You will contribute to cutting‑edge foundational AI research and advanced HPC capabilities.
We are a dedicated research lab for building, understanding, using, and risk-managing foundation models. Our mandate is to advance research, nurture the next generation of AI builders, and drive transformative contributions to a knowledge-driven economy.
As part of our team, you’ll have the opportunity to work on the core of cutting‑edge foundation model training, alongside world‑class researchers, data scientists, and engineers, tackling the most fundamental and impactful challenges in AI development. You will participate in the development of groundbreaking AI solutions that have the potential to reshape entire industries. Strategic and innovative problem‑solving skills will be instrumental in establishing MBZUAI as a global hub for high‑performance computing in deep learning, driving impactful discoveries that inspire the next generation of AI pioneers.
The Distributed ML Engineer will play a role at the forefront of optimizing performance for the machine learning software stacks, especially at training and inference, and support the team to develop new and cutting‑edge systems. The ideal candidate will have a strong background in parallel computing, and hands‑on experience in system level coding, debug methodologies, and large‑scale machine learning experience.
This position is eligible for visa sponsorship.