Get more replies from employers
Send a job-specific resume in minutes.
Amazon’s Annapurna Labs team, part of AWS, seeks a Software Development Engineer for AI/ML focused on AWS Neuron device acceleration. You will contribute to the Neuron SDK, optimize inference and training workloads on Inferentia/Trainium, and work across compiler, runtime, framework, and hardware teams.
The role emphasizes designing high-performance ML kernels, distributed inference support for PyTorch, and close collaboration with customers to enable efficient model deployment on AWS
The Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit that accelerates deep learning and GenAI workloads on Amazon’s custom machine learning accelerators, Inferentia and Trainium. The AWS Neuron SDK is the backbone for accelerating deep learning and GenAI workloads, offering an ML compiler, runtime, and application framework that integrates with popular ML frameworks like PyTorch and JAX to deliver top‑performance inference and training.
The Inference Enablement and Acceleration team focuses on running a wide range of models and supporting new architectures while maximizing performance on AWS’s custom ML accelerators. Working across the stack from PyTorch to the hardware‑software boundary, engineers build systematic infrastructure, innovate new methods, and create high‑performance kernels, ensuring every compute unit is fine‑tuned for optimal performance. This role offers the chance to work at the intersection of machine learning, high‑performance computing, and distributed architectures, shaping the future of AI acceleration technology.
Key responsibilities include leading distributed inference support for PyTorch in the Neuron SDK, tuning these models for highest performance on AWS Trainium and Inferentia silicon, developing low‑level optimizations, and collaborating with compiler, runtime, framework, and hardware teams to optimize machine learning workloads for the global customer base.
The Inference Enablement and Acceleration team fosters a builder’s culture where experimentation is encouraged, impact is measurable, and collaboration, technical ownership, and continuous learning are valued. The team provides mentorship, thorough but kind code reviews, and projects that develop engineering expertise, empowering members to handle complex tasks.
Learn more about Neuron:
https://awsdocs-neuron.readthedocs-hosted.com/en/latest/neuron-guide/neuron-cc/index.html
https://aws.amazon.com/machine-learning/neuron/
https://github.com/aws/aws-neuron-sdk
https://www.amazon.science/how-silicon-innovation-became-the-secret-sauce-behind-awss-success
Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.