ML Kernel Performance Engineering Manager

Amazon Web Services (AWS)

Cupertino (CA)

On-site

USD 212,700 - 287,700

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Annapurna Labs (U.S.) Inc. is seeking kernel engineers to optimize ML workloads on AWS accelerators.

You will design high-performance compute kernels for Neuron, analyze multi-generation performance, and implement compiler optimizations in collaboration with compiler, runtime, framework, and hardware teams. The role involves working at the hardware-software boundary, mentoring engineers, publishing research, and interfacing with customers to enable model acceleration on Inferentia and Trainium

Qualifications

  • 3+ years of engineering team management experience.
  • 7+ years of working directly within engineering teams experience.
  • 3+ years of designing or architecting multi-tier web services.
  • 8+ years of leading the definition and development of multi tier web services.
  • Experience partnering with product or program management teams.
  • Knowledge of engineering practices and patterns for the full software/hardware/networks development life cycle, including coding standards, code reviews, source control management, build processes, testing, certification, and livesite operations.
  • Experience in recruiting, mentoring/coaching and managing teams of Software Engineers.

Responsibilities

  • Design and implement high-performance compute kernels for ML operations, leveraging the Neuron architecture and programming models
  • Analyze and optimize kernel-level performance across multiple generations of Neuron hardware
  • Conduct detailed performance analysis using profiling tools to identify and resolve bottlenecks
  • Implement compiler optimizations such as fusion, sharding, tiling, and scheduling
  • Work directly with customers to enable and optimize their ML models on AWS accelerators
  • Collaborate across teams to develop innovative kernel optimization techniques

Skills

Engineering management
Engineering teams experience
System design
Web services architecture
Software lifecycle
Product collaboration

Job description

Annapurna Labs (U.S.) Inc. is seeking kernel engineers to optimize ML workloads on AWS accelerators.

You will design high-performance compute kernels for Neuron, analyze multi-generation performance, and implement compiler optimizations in collaboration with compiler, runtime, framework, and hardware teams. The role involves working at the hardware-software boundary, mentoring engineers, publishing research, and interfacing with customers to enable model acceleration on Inferentia and Trainium

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Kernel Performance Engineer for AI Accelerators
ML Kernel Performance Engineer for AI Accelerators

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
Health insurance
RSUs
401(k) matching
+1
ML Kernel Performance Engineer for Neuron Accelerators
ML Kernel Performance Engineer for Neuron Accelerators

Amazon • Cupertino (CA)

On-site
USD 140,000 - 210,000
Senior ML Kernel Optimization Engineer
Senior ML Kernel Optimization Engineer

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
Senior ML Kernel Performance Architect
Senior ML Kernel Performance Architect

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
Engineering Manager, ML Kernel Performance
Engineering Manager, ML Kernel Performance

Amazon • Cupertino (CA)

On-site
USD 212,700 - 287,700
Health insurance
401(k) matching
Parental leave
+2
Sr. ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs
Sr. ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs
ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs

Amazon • Cupertino (CA)

On-site
USD 140,000 - 210,000
AI/ML Inference Engineer for AWS Neuron
AI/ML Inference Engineer for AWS Neuron

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
Linux Driver & Runtime Engineer for ML Accelerator SDK
Linux Driver & Runtime Engineer for ML Accelerator SDK

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 143,000 - 195,000
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs

Amazon • Cupertino (CA)

On-site
USD 212,700 - 287,700
Health insurance
401(k) matching
Parental leave
+2