Senior AI/ML Software Engineer - High-Perf Inference

Amazon Web Services (AWS)

Seattle (WA)

On-site

USD 168,000 - 227,000

Full time

28 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Annapurna Labs (U.S.) Inc. in Seattle seeks a Senior Software Engineer to advance AWS Neuron’s AI accelerators. You will optimize deep learning models and work across compiler, runtime, framework, and hardware teams to push performance on Inferentia and Trainium.

You will design kernels, monitor system performance, and collaborate with customers to maximize efficiency of ML workloads on AWS accelerators.

Qualifications

  • Bachelor’s degree required.
  • 5+ years of professional software development experience.
  • Knowledge of Python and/or C++ programming.

Responsibilities

  • Lead distributed inference support for PyTorch in the Neuron SDK.
  • Tune models for highest performance on Trainium/Inferentia silicon.
  • Collaborate across compiler, runtime, framework, and hardware teams.
  • Design and implement high-performance ML kernels and features for Neuron.

Skills

Python programming
C++ programming
Performance profiling
Parallel computing
System design leadership

Education

Master’s degree in Computer Science or equivalent
Bachelor’s degree

Tools

PyTorch
CUDA
Triton
NCCL

Job description

Annapurna Labs (U.S.) Inc. in Seattle seeks a Senior Software Engineer to advance AWS Neuron’s AI accelerators. You will optimize deep learning models and work across compiler, runtime, framework, and hardware teams to push performance on Inferentia and Trainium.

You will design kernels, monitor system performance, and collaborate with customers to maximize efficiency of ML workloads on AWS accelerators.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior ML Systems Engineer - AI Inference on AWS Neuron
Senior ML Systems Engineer - AI Inference on AWS Neuron

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
AI/ML Inference Engineer for AWS Neuron
AI/ML Inference Engineer for AWS Neuron

Amazon • Cupertino (CA)

On-site
USD 165,000 - 224,000
Health insurance
401(k) matching
Paid time off
+2
ML Inference Engineer - AWS Neuron & GenAI
ML Inference Engineer - AWS Neuron & GenAI

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
AI/ML Systems Engineer - Neuron Inference
AI/ML Systems Engineer - Neuron Inference

Amazon • Seattle (WA)

On-site
USD 144,000 - 194,000
Senior AI/ML Software Engineer - Neuron Optimizations
Senior AI/ML Software Engineer - Neuron Optimizations

Annapurna Labs (U.S.) Inc. • Seattle (WA)

On-site
USD 180,000 - 230,000
Senior ML Kernel Performance Engineer - AI Accelerator
Senior ML Kernel Performance Engineer - AI Accelerator

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
Senior ML Compiler Engineer – Neuron
Senior ML Compiler Engineer – Neuron

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 168,000 - 227,000
Senior AI/ML Systems Engineer for Accelerator Optimization
Senior AI/ML Systems Engineer for Accelerator Optimization

Socket.dev • Seattle (WA)

On-site
USD 168,000 - 227,000
Senior Software Engineer — GenAI & ML Acceleration
Senior Software Engineer — GenAI & ML Acceleration

Amazon Web Services (AWS) • New York (NY)

On-site
USD 185,000 - 250,000
Health insurance
401(k) matching
Paid time off
+2
Senior AI/ML Systems Engineer - Disaggregated Inference
Senior AI/ML Systems Engineer - Disaggregated Inference

Energy Jobline ZR • Cupertino (CA)

On-site
USD 193,000 - 262,000
Health insurance
401(k) matching
Paid time off
+1