Senior ML Kernel Optimization Engineer

Amazon

Toronto

Hybrid

CAD 151,000 - 252,000

Full time

36 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
RRSP
DPSP
Paid time off

Job summary

Amazon's Annapurna Labs team is hiring engineers to design and optimize high-performance compute kernels for ML on Neuron. You will work across Neuron architecture generations, profiling and tuning for maximum throughput, and implement compiler optimizations while collaborating with customers to enable their models on AWS accelerators.

The role blends machine learning, high-performance computing, and distributed architectures in a startup-like environment, with mentorship and customer-facing

Qualifications

  • 5+ years of full software development life cycle experience.
  • Bachelor's degree in computer science or equivalent.
  • Expertise in accelerator architectures for ML/HPC (GPUs/FPGA/ASIC).
  • Experience with CUDA, OpenCL, ROCm and GPU backends.
  • Proficiency with NVIDIA PTX and GPU ISA concepts.
  • Experience developing high-performance HPC libraries.
  • Strong low-level optimization for GPUs and ML frameworks.

Responsibilities

  • Design and implement high-performance compute kernels for ML on Neuron.
  • Optimize kernel performance across multiple Neuron generations.
  • Perform detailed profiling to identify bottlenecks and improve throughput.
  • Apply compiler optimizations: fusion, tiling, scheduling.
  • Collaborate with customers to optimize models on AWS accelerators.
  • Work cross-functionally to develop innovative kernel techniques.

Skills

SDLC experience
Accelerator architectures
GPU kernel optimization
NVIDIA PTX
High-performance libraries
GPU memory optimization
LLVM/MLIR backends
PyTorch / TensorFlow GPUs
Parallel programming
GPU backends knowledge

Education

Bachelor's degree

Tools

CUDA
OpenCL
ROCm
PyTorch
TensorFlow
LLVM

Job description

Amazon's Annapurna Labs team is hiring engineers to design and optimize high-performance compute kernels for ML on Neuron. You will work across Neuron architecture generations, profiling and tuning for maximum throughput, and implement compiler optimizations while collaborating with customers to enable their models on AWS accelerators.

The role blends machine learning, high-performance computing, and distributed architectures in a startup-like environment, with mentorship and customer-facing

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Engineering Manager — ML Kernel Performance
Senior Engineering Manager — ML Kernel Performance

Amazon • Toronto

On-site
CAD 171,000 - 286,000
ML Kernel Performance Eng. Manager, Accelerators
ML Kernel Performance Eng. Manager, Accelerators

Socket.dev • Toronto

On-site
CAD 171,000 - 286,000
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs

Amazon • Toronto

On-site
CAD 171,000 - 286,000
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs

Socket.dev • Toronto

On-site
CAD 171,000 - 286,000
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs

Amazon Web Services (AWS) • Toronto

On-site
CAD 171,000 - 286,000
Senior ML Kernel Performance Engineer
Senior ML Kernel Performance Engineer

Amazon • Toronto

Hybrid
CAD 151,000 - 252,000
Health insurance
RRSP
DPSP
+1
Kernel Engineer
Kernel Engineer

Acceler8 Talent • Canada

Remote
CAD 109,000 - 193,000
Sr. Software Development Manager - Compiler, AWS Neuron, Annapurna Labs
Sr. Software Development Manager - Compiler, AWS Neuron, Annapurna Labs

Amazon Web Services (AWS) • Toronto

On-site
CAD 214,000 - 358,000
Senior ML Infra Architect: High-Perf GPU & Scale
Senior ML Infra Architect: High-Perf GPU & Scale

Ellison Institute of Technology • Town of Oxford

On-site
CAD 110,000 - 170,000
Enhanced holiday pay
Pension
Life Assurance
+6
Senior ML Engineer - AI-Driven Optimization Systems
Senior ML Engineer - AI-Driven Optimization Systems

ServiceNow, Inc. • Toronto

Remote
CAD 120,000 - 190,000