Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
microTECH Global Limited is seeking a world-class compiler and performance optimization expert to lead the Triton compiler and kernel framework for NPUs. You will drive AI workloads, implement advanced Triton kernels, and optimize memory, scheduling, and end-to-end performance with PyTorch and XLA.
The role demands a strong background in AI compilation, NPU programming, and systems performance engineering, with hands-on kernel development and mentoring responsibilities. PhD is a plus.
We are looking for a world-classcompiler and performance optimization expert to join our deep learning infrastructure team. You will take ownership of building a high-performance Triton compiler and kernel optimization framework, driving the next generation of AI workloads on NPUs. This is a highly technical role that sits at the intersection ofAI compilation, NPU programming, and system performance engineering.