Senior ML Kernel Optimizer for AWS Neuron

Amazon Web Services (AWS)

Cupertino (CA)

On-site

USD 193,000 - 262,000

Full time

5 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Annapurna Labs (U.S.) Inc. seeks experienced kernel engineers to optimize ML workloads on AWS accelerators. You will design high-performance compute kernels for ML ops and work across compiler, runtime, framework, and hardware teams to maximize Neuron-based performance.

The role emphasizes low-level optimization, system architecture, and collaboration with customers to enable model acceleration on Inferentia and Trainium. Cupertino, CA-based team location with a strong learning culture.

Qualifications

  • 5+ years of non-internship professional software development experience.
  • 5+ years of programming with at least one software programming language experience.
  • 5+ years of leading design or architecture (design patterns, reliability and scaling) of new and existing systems experience.
  • 5+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience.
  • Experience as a mentor, tech lead or leading an engineering team.

Responsibilities

  • Design and implement high-performance compute kernels for ML operations, leveraging the Neuron architecture and programming models
  • Analyze and optimize kernel-level performance across multiple generations of Neuron hardware
  • Conduct detailed performance analysis using profiling tools to identify and resolve bottlenecks
  • Implement compiler optimizations such as fusion, sharding, tiling, and scheduling
  • Work directly with customers to enable and optimize their ML models on AWS accelerators
  • Collaborate across teams to develop innovative kernel optimization techniques

Skills

Software development
Programming languages
Leadership
Mentor / tech lead
SDLC / CI/CD
GPU kernel optimization

Education

Bachelor's degree in computer science or equivalent

Tools

CUDA
NVIDIA PTX
LLVM/MLIR

Job description

Annapurna Labs (U.S.) Inc. seeks experienced kernel engineers to optimize ML workloads on AWS accelerators. You will design high-performance compute kernels for ML ops and work across compiler, runtime, framework, and hardware teams to maximize Neuron-based performance.

The role emphasizes low-level optimization, system architecture, and collaboration with customers to enable model acceleration on Inferentia and Trainium. Cupertino, CA-based team location with a strong learning culture.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

ML Kernel Performance Engineer for Neuron Accelerators
ML Kernel Performance Engineer for Neuron Accelerators

Amazon • Cupertino (CA)

On-site
USD 140,000 - 210,000
Senior ML Kernel Performance Engineer - AI Accelerator
Senior ML Kernel Performance Engineer - AI Accelerator

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
ML Inference Engineer - AWS Neuron & GenAI
ML Inference Engineer - AWS Neuron & GenAI

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
Senior AI/ML Inference Engineer (Neuron)
Senior AI/ML Inference Engineer (Neuron)

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
Health insurance
401(k) matching
Paid time off
Senior ML Systems Engineer - AI Inference on AWS Neuron
Senior ML Systems Engineer - AI Inference on AWS Neuron

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
AI/ML Systems Engineer - Neuron Inference
AI/ML Systems Engineer - Neuron Inference

Amazon • Seattle (WA)

On-site
USD 144,000 - 194,000
Senior ML Compiler Engineer – Neuron
Senior ML Compiler Engineer – Neuron

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 168,000 - 227,000
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs

Amazon • Cupertino (CA)

On-site
USD 140,000 - 210,000
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs

Annapurna Labs (U.S.) Inc. - D63 • Cupertino (CA)

On-site
USD 180,000 - 240,000
Senior AI/ML Software Engineer - High-Perf Inference
Senior AI/ML Software Engineer - High-Perf Inference

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 168,000 - 227,000