ML Kernel Performance Eng. Manager, Accelerators

Socket.dev

Toronto

On-site

CAD 171,000 - 286,000

Full time

11 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Amazon in Toronto is seeking a senior ML kernel engineer to design and optimize high-performance compute kernels for AWS Neuron on Inferentia and Trainium accelerators. You will work across multiple generations, analyze bottlenecks, and collaborate with customers to enable optimized ML models.

The role combines ML, HPC, and distributed architectures, offering opportunities to mentor engineers, publish research, and drive kernel innovations for broad customer impact.

Qualifications

  • 3+ years of engineering team management experience.
  • 7+ years of working directly within engineering teams experience.
  • 3+ years of designing or architecting (design patterns, reliability and scaling) of new and existing systems experience.
  • 8+ years of leading the definition and development of multi tier web services experience.
  • Knowledge of engineering practices and patterns for the full software/hardware/networks development life cycle, including coding standards, code reviews, source control management, build processes, testing, certification, and livesite operations.
  • Experience partnering with product or program management teams.

Responsibilities

  • Design and implement high-performance compute kernels for ML operations, leveraging the Neuron architecture and programming models.
  • Analyze and optimize kernel-level performance across multiple generations of Neuron hardware.
  • Conduct detailed performance analysis using profiling tools to identify and resolve bottlenecks.
  • Implement compiler optimizations such as fusion, sharding, tiling, and scheduling.
  • Work directly with customers to enable and optimize their ML models on AWS accelerators.
  • Collaborate across teams to develop innovative kernel optimization techniques.

Skills

Team management
Engineering experience
Systems design
Web services
Software lifecycle
Product collaboration
Requirements gathering
People management

Job description

Amazon in Toronto is seeking a senior ML kernel engineer to design and optimize high-performance compute kernels for AWS Neuron on Inferentia and Trainium accelerators. You will work across multiple generations, analyze bottlenecks, and collaborate with customers to enable optimized ML models.

The role combines ML, HPC, and distributed architectures, offering opportunities to mentor engineers, publish research, and drive kernel innovations for broad customer impact.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Engineering Manager — ML Kernel Performance
Senior Engineering Manager — ML Kernel Performance

Amazon • Toronto

On-site
CAD 171,000 - 286,000
Senior ML Kernel Performance Engineer
Senior ML Kernel Performance Engineer

Amazon Web Services (AWS) • Toronto

On-site
CAD 150,000 - 252,000
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs

Socket.dev • Toronto

On-site
CAD 171,000 - 286,000
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs

United States Digital Space LLC • Toronto

On-site
CAD 171,000 - 286,000
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs

Amazon • Toronto

On-site
CAD 171,000 - 286,000
Sr. Software Development Manager - Compiler, AWS Neuron, Annapurna Labs
Sr. Software Development Manager - Compiler, AWS Neuron, Annapurna Labs

Amazon Web Services (AWS) • Toronto

On-site
CAD 214,000 - 358,000
Kernel Engineer
Kernel Engineer

Acceler8 Talent • Canada

Remote
CAD 109,000 - 193,000
Senior ML Performance Engineer - Distributed Training
Senior ML Performance Engineer - Distributed Training

Veeda AI • Toronto

On-site
CAD 120,000 - 180,000
Senior ML Compiler Engineer - HW/SW Co-Design
Senior ML Compiler Engineer - HW/SW Co-Design

Qualcomm • Markham

On-site
CAD 120,000 - 180,000
Senior ML Infra Architect: High-Perf GPU & Scale
Senior ML Infra Architect: High-Perf GPU & Scale

Ellison Institute of Technology • Town of Oxford

On-site
CAD 110,000 - 170,000
Enhanced holiday pay
Pension
Life Assurance
+6