ML Kernel Performance Engineer for AI Accelerators

Amazon Web Services (AWS)

Cupertino (CA)

On-site

USD 165,200 - 223,600

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
RSUs
401(k) matching
Paid time off

Job summary

Annapurna Labs (U.S.) Inc. in Cupertino, CA, is seeking kernel engineers to optimize ML workloads on AWS accelerators.

You will design and implement high-performance kernels, analyze kernel-level performance, and collaborate across hardware, software, and ML teams to push the performance of AWS Neuron for Inferentia and Trainium. Join a team that blends deep hardware knowledge with ML expertise to advance AI acceleration, publish cutting-edge research, and mentor engineers in a startup-like,

Qualifications

  • 3+ years of non-internship professional software development experience.
  • 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems.
  • Experience programming with at least one software programming language.
  • Bachelor's degree in computer science or equivalent.

Responsibilities

  • Design and implement high-performance compute kernels for ML operations, leveraging the Neuron architecture and programming models.
  • Analyze and optimize kernel-level performance across multiple generations of Neuron hardware.
  • Conduct detailed performance analysis using profiling tools to identify and resolve bottlenecks.
  • Implement compiler optimizations such as fusion, sharding, tiling, and scheduling.
  • Work directly with customers to enable and optimize their ML models on AWS accelerators.
  • Collaborate across teams to develop innovative kernel optimization techniques.

Skills

Software development
Design/Architecture
Programming languages

Education

Bachelor's degree in Computer Science or equivalent

Job description

Annapurna Labs (U.S.) Inc. in Cupertino, CA, is seeking kernel engineers to optimize ML workloads on AWS accelerators.

You will design and implement high-performance kernels, analyze kernel-level performance, and collaborate across hardware, software, and ML teams to push the performance of AWS Neuron for Inferentia and Trainium. Join a team that blends deep hardware knowledge with ML expertise to advance AI acceleration, publish cutting-edge research, and mentor engineers in a startup-like,

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Kernel Performance Engineering Manager
ML Kernel Performance Engineering Manager

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 212,000 - 288,000
ML Kernel Performance Engineer for Neuron Accelerators
ML Kernel Performance Engineer for Neuron Accelerators

Amazon • Cupertino (CA)

On-site
USD 140,000 - 210,000
Senior ML Kernel Optimization Engineer
Senior ML Kernel Optimization Engineer

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
Senior ML Kernel Performance Architect
Senior ML Kernel Performance Architect

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
Engineering Manager, ML Kernel Performance
Engineering Manager, ML Kernel Performance

Amazon • Cupertino (CA)

On-site
USD 212,700 - 287,700
Health insurance
401(k) matching
Parental leave
+2
AI/ML Inference Engineer for AWS Neuron
AI/ML Inference Engineer for AWS Neuron

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs

Amazon • Cupertino (CA)

On-site
USD 140,000 - 210,000
Sr. ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs
Sr. ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
Linux Driver & Runtime Engineer for ML Accelerator SDK
Linux Driver & Runtime Engineer for ML Accelerator SDK

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 143,000 - 195,000
Applied Scientist II — ML Systems for AI Accelerators
Applied Scientist II — ML Systems for AI Accelerators

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 171,000 - 223,000