Senior AI/ML Systems Engineer for Accelerator Optimization

Socket.dev

Seattle (WA)

On-site

USD 168,000 - 227,000

Full time

3 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Amazon, the maker of AWS Neuron, seeks a Senior Software Engineer to optimize AI model inference on custom accelerators like Inferentia and Trainium. You will work across compiler, runtime, and hardware teams to push the performance of Generative AI workloads and ML systems at scale.

You will design high-performance kernels, profiling and tuning ML workloads, and collaborating with customers and open-source communities to deliver optimized solutions on AWS accelerators.

Qualifications

  • Bachelor's degree in CS or related field required.
  • 5+ years of professional software development experience.
  • Proficiency in Python and/or C++.
  • 5+ years of design/architecture leadership for systems.
  • Experience debugging, profiling, and optimizing large-scale systems.
  • Strong knowledge of memory management and parallel computing.
  • Experience owning a performance optimization roadmap.

Responsibilities

  • Design, develop, and optimize ML models on custom AI accelerators.
  • Lead in distributed inference support for PyTorch in Neuron SDK.
  • Profile and optimize performance across Neuron hardware generations.
  • Build infrastructure to onboard multiple models with diverse architectures.
  • Develop high-performance kernels for ML operations on Neuron hardware.
  • Collaborate across teams to enable and optimize customer ML workloads.

Skills

Python programming
C++ programming
System performance
Debugging and profiling
Parallel computing

Education

Bachelor's degree
Master’s degree

Tools

PyTorch
CUDA
Triton
NCCL

Job description

Amazon, the maker of AWS Neuron, seeks a Senior Software Engineer to optimize AI model inference on custom accelerators like Inferentia and Trainium. You will work across compiler, runtime, and hardware teams to push the performance of Generative AI workloads and ML systems at scale.

You will design high-performance kernels, profiling and tuning ML workloads, and collaborating with customers and open-source communities to deliver optimized solutions on AWS accelerators.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI/ML Software Engineer - Neuron Optimizations
Senior AI/ML Software Engineer - Neuron Optimizations

Annapurna Labs (U.S.) Inc. • Seattle (WA)

On-site
USD 180,000 - 230,000
Senior ML Accelerator Runtime Engineer
Senior ML Accelerator Runtime Engineer

Amazon • Seattle (WA)

On-site
USD 143,700 - 194,400
Health insurance
Dental
Vision
+4
Senior ML Inference Systems Engineer - Custom Accelerator
Senior ML Inference Systems Engineer - Custom Accelerator

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
Health insurance
RSUs
401(k) matching
Senior ML Compiler Engineer – Build Next‑Gen ML on AI Accelerators
Senior ML Compiler Engineer – Build Next‑Gen ML on AI Accelerators

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
Health benefits
401(k) matching
Parental leave
Sr. Software Engineer- AI/ML, AWS Neuron
Sr. Software Engineer- AI/ML, AWS Neuron

Annapurna Labs (U.S.) Inc. • Seattle (WA)

On-site
USD 180,000 - 230,000
AI/ML Systems Engineer - Neuron Inference
AI/ML Systems Engineer - Neuron Inference

Amazon • Seattle (WA)

On-site
USD 144,000 - 194,000
ML Compiler Engineer for Custom ML Accelerators
ML Compiler Engineer for Custom ML Accelerators

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 180,000 - 240,000
Senior ML Systems Engineer - AI Inference on AWS Neuron
Senior ML Systems Engineer - AI Inference on AWS Neuron

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
Sr. Software Engineer- AI/ML, AWS Neuron
Sr. Software Engineer- AI/ML, AWS Neuron

Socket.dev • Seattle (WA)

On-site
USD 168,000 - 227,000
ML Kernel Performance Engineer for Neuron Accelerators
ML Kernel Performance Engineer for Neuron Accelerators

Amazon • Cupertino (CA)

On-site
USD 140,000 - 210,000