Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Itlearn360 in Cupertino is seeking an experienced distributed AI/ML systems engineer to develop collective operations enabling scale across multiple accelerators and servers.
Most of our stack is C/C++ and low level; you will work with Linux kernels, HPC interconnects, embedded systems, and high‑speed networking to deliver scalable solutions for large models and clusters.
We are seeking an experienced engineer to work on distributed AI/ML systems and develop collective operations that enable AI to scale across multiple accelerators and servers.
Most of our stack is C/C++ and relatively low level, so solid knowledge of Linux, kernels, and performant code is important. Experience with embedded systems, high‑speed networking, and HPC interconnects is valued highly.
If you like solving hard problems, working with HPC and ML customers, iterating fast, and delivering meaningful solutions at scale, this role is for you. You will work on features for the largest clusters, customers, and AI models in a fast‑moving environment while collaborating with infrastructure, hardware, RTL engineers, scientists, and architects.