Engineering Manager, ML Kernel Performance

Amazon Web Services (AWS)

Cupertino (CA)

On-site

USD 213,000 - 288,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
RSUs
401(k) matching
Paid time off
Parental leave

Job summary

Annapurna Labs (U.S.) Inc. in Cupertino, CA is seeking a senior kernel engineer to design and optimize high-performance ML compute kernels for AWS accelerators. You will work across compiler, runtime, and hardware teams to accelerate ML workloads and improve performance.

You will analyze bottlenecks, apply advanced kernel optimizations, and collaborate with customers to enable their models on Neuron hardware. This role combines software, hardware, and ML systems in a startup-like environment.

Qualifications

  • 3+ years of engineering team management experience.
  • 7+ years of working directly within engineering teams experience.
  • 3+ years of designing or architecting multi-tier web services.
  • 8+ years of leading the definition and development of multi tier web services.

Responsibilities

  • Design and implement high-performance compute kernels for ML operations using the Neuron architecture.
  • Analyze and optimize kernel-level performance across Neuron hardware generations.
  • Conduct detailed performance analysis with profiling tools to identify bottlenecks.
  • Implement compiler optimizations such as fusion, sharding, tiling, and scheduling.
  • Work directly with customers to enable and optimize ML models on AWS accelerators.
  • Collaborate across teams to develop innovative kernel optimization techniques.

Skills

Engineering management
Engineering teams
Low-level optimization
System architecture
ML model acceleration
Profiling tools

Job description

Annapurna Labs (U.S.) Inc. in Cupertino, CA is seeking a senior kernel engineer to design and optimize high-performance ML compute kernels for AWS accelerators. You will work across compiler, runtime, and hardware teams to accelerate ML workloads and improve performance.

You will analyze bottlenecks, apply advanced kernel optimizations, and collaborate with customers to enable their models on Neuron hardware. This role combines software, hardware, and ML systems in a startup-like environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Engineering Manager, ML Kernel Performance
Engineering Manager, ML Kernel Performance

Amazon • Cupertino (CA)

On-site
USD 212,700 - 287,700
Health insurance
401(k) matching
Parental leave
+2
ML Kernel Performance Engineer for Neuron
ML Kernel Performance Engineer for Neuron

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
ML Kernel Performance Engineer for Neuron Accelerators
ML Kernel Performance Engineer for Neuron Accelerators

Amazon • Cupertino (CA)

On-site
USD 140,000 - 210,000
ML Inference Engineer - AWS Neuron & GenAI
ML Inference Engineer - AWS Neuron & GenAI

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs

Amazon • Cupertino (CA)

On-site
USD 212,700 - 287,700
Health insurance
401(k) matching
Parental leave
+2
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs

Amazon • Cupertino (CA)

On-site
USD 140,000 - 210,000
C/C++ ML Acceleration HW/SW Engineer
C/C++ ML Acceleration HW/SW Engineer

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
Senior Hardware Engineer, ML Acceleration & SoC
Senior Hardware Engineer, ML Acceleration & SoC

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 183,000 - 248,000
Health insurance
401(k) matching
Paid time off
+1
Senior ML Systems Engineer - AI Inference on AWS Neuron
Senior ML Systems Engineer - AI Inference on AWS Neuron

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
Embedded C/C++ Engineer for ML Accelerator Systems
Embedded C/C++ Engineer for ML Accelerator Systems

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 144,000 - 194,000
Health insurance
401(k) matching
Paid time off
+2