ML Kernel Performance Engineer for Neuron Accelerators

Amazon

Cupertino (CA)

On-site

USD 140,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Amazon's Annapurna Labs in the AWS Neuron group seeks a dedicated ML Kernel Performance Engineer to design and optimize high-performance compute kernels for ML workloads on Inferentia and Trainium. You will push kernel optimization boundaries across generations and collaborate with cross-functional teams.

This role partners directly with customers to tune models for AWS accelerators, applies compiler optimizations, and participates in design reviews and code quality activities in a startup-like

Qualifications

  • 3+ years of non-internship professional software development experience.
  • 2+ years of non-internship design or architecture experience for new and existing systems.
  • Experience programming with at least one software programming language.
  • 3+ years of full software development life cycle experience, including coding standards, reviews, and operations.
  • Bachelor’s degree in computer science or equivalent.

Responsibilities

  • Design and implement high-performance compute kernels for ML operations leveraging the Neuron architecture.
  • Analyze and optimize kernel-level performance across multiple generations of Neuron hardware.
  • Conduct detailed performance profiling to identify bottlenecks and optimize them.
  • Apply compiler optimizations such as fusion, tiling, and scheduling.
  • Work directly with customers to enable and optimize ML models on AWS accelerators.

Skills

Software development
System design
Programming language

Education

Bachelor's degree in Computer Science

Job description

Amazon's Annapurna Labs in the AWS Neuron group seeks a dedicated ML Kernel Performance Engineer to design and optimize high-performance compute kernels for ML workloads on Inferentia and Trainium. You will push kernel optimization boundaries across generations and collaborate with cross-functional teams.

This role partners directly with customers to tune models for AWS accelerators, applies compiler optimizations, and participates in design reviews and code quality activities in a startup-like

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Kernel Performance Engineering Manager
ML Kernel Performance Engineering Manager

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 212,000 - 288,000
ML Kernel Performance Engineer for AI Accelerators
ML Kernel Performance Engineer for AI Accelerators

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
Health insurance
RSUs
401(k) matching
+1
Senior ML Kernel Performance Architect
Senior ML Kernel Performance Architect

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
Senior ML Kernel Optimization Engineer
Senior ML Kernel Optimization Engineer

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
Engineering Manager, ML Kernel Performance
Engineering Manager, ML Kernel Performance

Amazon • Cupertino (CA)

On-site
USD 212,700 - 287,700
Health insurance
401(k) matching
Parental leave
+2
Sr. ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs
Sr. ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs

Amazon • Cupertino (CA)

On-site
USD 140,000 - 210,000
AI/ML Inference Engineer for AWS Neuron
AI/ML Inference Engineer for AWS Neuron

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs

Amazon • Cupertino (CA)

On-site
USD 212,700 - 287,700
Health insurance
401(k) matching
Parental leave
+2
Senior Neuron Runtime Software Engineer for ML Accelerators
Senior Neuron Runtime Software Engineer for ML Accelerators

Amazon • Cupertino (CA)

On-site
USD 150,000 - 210,000
Health insurance (medical, dental, and
Vision and prescription coverage