ML Systems Software Development Engineer Intern, Annapurna Labs - 2027

Socket.dev

Toronto

On-site

CAD 86,000 - 116,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Amazon in Toronto, Ontario is seeking a Software Development Engineer Intern to help bring up and optimize state-of-the-art machine learning models for peak performance on AWS Trainium. You will build profiling tools and contribute to industry-leading developer workflows.

You will join the Toronto Neuron team (Annapurna Labs) and work with mentors across ML and systems engineering to tackle real customer challenges, gaining hands-on experience with cutting-edge silicon, compilers, and runtimes.

Qualifications

  • Currently enrolled in a Bachelor's degree program in Computer Science, Computer Engineering, Electrical Engineering, or related field
  • Proficient in Python, C, and/or C++
  • Strong interest in performance engineering, kernel development, or ML systems technologies

Responsibilities

  • Join the Toronto Neuron team to bring up and optimize state-of-the-art machine learning models for peak performance on AWS Trainium
  • Build and improve profiling tools to help engineers identify bottlenecks and fix issues
  • Work on kernel/low-level optimizations and developer tooling as part of a cross-functional team
  • Collaborate with mentors across machine learning and systems engineering to solve real customer problems

Skills

Python
C
C++
Performance engineering
Kernel development

Education

Bachelor's degree in Computer Science/Computer Engineering/Electrical Engineering or related field

Tools

LLVM
MLIR
XLA
TVM

Job description

At Amazon, our tech isn't just a tool, it's your playground. Our engineers work on scalable systems, cloud services, and customer-facing products that operate at global scale. This is an environment where you learn by building. Whether you're creating cloud-native solutions, optimizing machine learning models, or building products that drive experiences for millions of customers around the globe, your work reaches the world, fast. You'll take ownership early, collaborate across teams and disciplines, and grow your skills as you tackle new problems. There's no single path forward here, with a chance to explore different directions and shape your journey as you go. This is where ambition meets opportunity and, where the impact of what you build helps define what comes next, for customers and for you.

What's in it for you? An internship at Amazon means real responsibility from day one. You'll build, test, and learn alongside people who want to see you succeed, while making an impact that reaches far beyond campus.

Key job responsibilities

Join the Toronto Neuron team (Annapurna Labs) to bring up and optimize state-of-the-art machine learning models for peak performance on AWS Trainium and build industry-leading profiling tools that help engineers identify and fix bottlenecks.

The Toronto Neuron team is hiring 2027 interns in two focus areas:
  • Frontier Model Performance team. Take newly released machine learning models from first bring-up to peak performance on current and next-generation AWS Trainium silicon. You could write and tune kernels, optimize sharding and model execution, build benchmark and measurement infrastructure, or tackle architectural bottlenecks that shape future Trainium designs. The best ideas do not stay trapped in one model. The team turns them into reusable Neuron components and optimization techniques for flagship open-source and customer workloads.
  • Developer Experience (DevEx) team. Make AWS Trainium performance visible. The team owns Neuron Explorer and builds state-of-the‑art profiling, debugging, and analysis tools that let engineers see how efficiently models and kernels use the hardware, pinpoint bottlenecks, and decide what to optimize next. You could build low-level performance data collection, analysis engines, interactive visualizations, or developer workflows. The goal is to cut through complex execution data so customers can debug faster and reach higher performance on Trainium.

Interns will join one of these teams based on their interests and experience. In either role, you will solve real customer and engineering problems with mentorship from experienced machine learning and systems engineers.

Internship Options
  • 12-16 month internship (starts May 2027)
  • 3-4 month internship (starts January 2027, May 2027, September 2027)
About the team

Annapurna Labs designs the custom silicon at Amazon - including Graviton (our server processors), Trainium and Inferentia (our machine learning training and inference accelerators), and the Nitro system that powers modern EC2. We own the full stack, from silicon through the software that makes it work (compilers, runtimes, drivers, and the frameworks that let customers run workloads on our hardware). As a Software Development Engineer Intern, you'll write the software that turns custom silicon into products used by millions.

Basic Qualifications
  • Currently enrolled in a Bachelor's degree program or higher in Computer Science, Computer Engineering, Electrical Engineering, or a related field
  • Programming experience through coursework, research, or a previous internship using Python, C, and/or C++
  • Strong interest and academic, research, or project experience in at least two of the following areas: 1/ Performance engineering, profiling, benchmarking, or low-level systems optimization. 2/ Kernel development, parallel programming, or computer architecture. 3/ Developer tooling, including profilers, debuggers, diagnostics, or visualization. 4/ Data structures and algorithms. 5/ Machine learning frameworks and models, including PyTorch or JAX. 6/ Compiler or ML systems technologies such as LLVM, MLIR, XLA, or TVM
Preferred Qualifications
  • Experience bringing up and optimizing machine learning models
  • Experience writing or optimizing kernels for GPUs, ML accelerators, or FPGAs
  • Experience using performance analysis tools or building developer tools
  • Previous technical internship or relevant research experience
  • Experience with full stack development, including front-end technologies such as TypeScript and React, and back-end services (experience with Go is a plus)
  • Ability to explain technical challenges and solutions clearly
  • Ability to independently work through ambiguous or undefined problems and think abstractly
  • Experience in developing agentic workflows for work automation

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

The starting pay for this position is listed below. In addition, Amazon offers basic life & AD&D insurance, paid time off, and other resources to improve health and well-being. We thank all applicants for their interest, however only those interviewed will be advised as to hiring status.

Toronto, ON, CAN - 100,810.00 CAD Annually

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr. Software Development Manager - Compiler, AWS Neuron, Annapurna Labs
Sr. Software Development Manager - Compiler, AWS Neuron, Annapurna Labs

Amazon Web Services (AWS) • Toronto

On-site
CAD 214,000 - 358,000
Software Development Engineer, Neuron Explorer
Software Development Engineer, Neuron Explorer

Socket.dev • Toronto

On-site
CAD 115,000 - 192,000
Senior ML Kernel Performance Engineer
Senior ML Kernel Performance Engineer

Amazon Web Services (AWS) • Toronto

On-site
CAD 151,000 - 252,000
Health insurance
RRSP and benefits
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs

Amazon Web Services (AWS) • Toronto

On-site
CAD 171,000 - 286,000
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs

Socket.dev • Toronto

On-site
CAD 171,000 - 286,000
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs

Amazon • Toronto

On-site
CAD 171,000 - 286,000
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs

Amazon Web Services (AWS) • Toronto

On-site
CAD 115,000 - 192,000
Health insurance (medical, dental, vis
RRSP
Deferred Profit Sharing Plan (DPSP)
+1
Senior ML Kernel Performance Engineer
Senior ML Kernel Performance Engineer

Amazon • Toronto

On-site
CAD 151,000 - 252,000
ML Systems Engineer Intern: Performance & Profiling
ML Systems Engineer Intern: Performance & Profiling

Socket.dev • Toronto

On-site
CAD 86,000 - 116,000
Software Development Engineer, Early Career - 2026
Software Development Engineer, Early Career - 2026

Amazon • Vancouver

On-site
CAD 90,000 - 150,000
Health insurance
RRSP
Paid time off