Senior ML Kernel Performance Engineer - AI Accelerator

Amazon

Cupertino (CA)

On-site

USD 193,000 - 262,000

Full time

2 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Amazon's Annapurna Labs seeks a Sr. ML Kernel Performance Engineer to advance AWS Neuron on Inferentia and Trainium. You will design high-performance kernels and optimize across generations, collaborating with compiler, runtime, framework, and hardware teams to push ML throughput.

This role offers exposure to cutting-edge AI acceleration and opportunities to mentor engineers while delivering production-grade kernel optimizations that scale with customer needs.

Qualifications

  • 5+ years of non-internship professional software development experience.
  • 5+ years of programming with at least one software programming language.
  • 5+ years of leading design or architecture of new and existing systems.
  • 5+ years of full software development life cycle, including coding standards and reviews.
  • Experience as a mentor, tech lead or leading an engineering team.
  • 6+ years of full software development experience.

Responsibilities

  • Design and implement high-performance compute kernels for ML operations, leveraging the Neuron architecture and programming models.
  • Analyze and optimize kernel-level performance across multiple generations of Neuron hardware.
  • Conduct detailed performance analysis using profiling tools to identify bottlenecks.
  • Implement compiler optimizations such as fusion, sharding, tiling, and scheduling.
  • Work directly with customers to enable and optimize their ML models on AWS accelerators.
  • Collaborate across teams to develop innovative kernel optimization techniques.

Skills

Kernel optimization
Performance analysis
Leadership/mentoring
ML acceleration
GPU architectures

Education

Bachelor's degree in computer science or equivalent

Tools

CUDA
NVIDIA PTX
Triton
OpenCL
SYCL
ROCm

Job description

Amazon's Annapurna Labs seeks a Sr. ML Kernel Performance Engineer to advance AWS Neuron on Inferentia and Trainium. You will design high-performance kernels and optimize across generations, collaborating with compiler, runtime, framework, and hardware teams to push ML throughput.

This role offers exposure to cutting-edge AI acceleration and opportunities to mentor engineers while delivering production-grade kernel optimizations that scale with customer needs.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Kernel Performance Engineer for Neuron Accelerators
ML Kernel Performance Engineer for Neuron Accelerators

Amazon • Cupertino (CA)

On-site
USD 140,000 - 210,000
ML Kernel Performance Engineer for Neuron
ML Kernel Performance Engineer for Neuron

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
Senior AI/ML Software Engineer - Neuron Optimizations
Senior AI/ML Software Engineer - Neuron Optimizations

Annapurna Labs (U.S.) Inc. • Seattle (WA)

On-site
USD 180,000 - 230,000
Senior AI/ML Systems Engineer for Accelerator Optimization
Senior AI/ML Systems Engineer for Accelerator Optimization

Socket.dev • Seattle (WA)

On-site
USD 168,000 - 227,000
Senior AI/ML Software Engineer - High-Perf Inference
Senior AI/ML Software Engineer - High-Perf Inference

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 168,000 - 227,000
AI/ML Systems Engineer - Neuron Inference
AI/ML Systems Engineer - Neuron Inference

Amazon • Seattle (WA)

On-site
USD 144,000 - 194,000
AI/ML Inference Engineer for AWS Neuron
AI/ML Inference Engineer for AWS Neuron

Amazon • Cupertino (CA)

On-site
USD 165,000 - 224,000
Health insurance
401(k) matching
Paid time off
+2
Senior ML Systems Engineer - AI Inference on AWS Neuron
Senior ML Systems Engineer - AI Inference on AWS Neuron

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs

Annapurna Labs (U.S.) Inc. - D63 • Cupertino (CA)

On-site
USD 180,000 - 240,000
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs

Amazon • Cupertino (CA)

On-site
USD 140,000 - 210,000