Senior ML Accelerator Runtime Engineer

Amazon

Seattle (WA)

On-site

USD 143,700 - 194,400

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Health insurance
Dental
Vision
RSUs
401(k) matching
Paid time off
Parental leave

Job summary

Amazon is seeking a Software Development Engineer for the AWS Neuron Runtime Team to build and maintain high‑performance runtime libraries and drivers for AI accelerators such as Inferentia and Trainium. You will design, develop, and deploy Neuron Runtime and other Neuron components, and contribute to profiling to optimize AI workloads across devices.

You will work with cross‑functional teams to enable performance insights, own end‑to‑end services, and participate in post‑incident reviews while

Qualifications

  • 3+ years of professional software development experience.
  • 2+ years of design or architecture experience focusing on reliability and scaling.
  • Experience programming in at least one software programming language.

Responsibilities

  • Design, develop, and deploy high‑performance runtime libraries and drivers for AI accelerators.
  • Develop and maintain the Neuron Runtime, ensuring scalability, reliability, and usability across services.
  • Collaborate with cross‑functional teams to generate key information in a C++ compiler that enables customers to understand and optimize performance.
  • Drive innovation to support multiple machine‑learning frameworks (PyTorch, JAX, XLA) in the profiler.
  • Own services end‑to‑end, including deployment, monitoring, alarming, on‑call responsibilities, and post‑incident reviews.
  • Work with executive leadership and senior technical leaders to define product directions and deliver solutions at scale.

Skills

Software development
System design
Programming language

Education

Bachelor's degree in computer science or equivalent

Job description

Amazon is seeking a Software Development Engineer for the AWS Neuron Runtime Team to build and maintain high‑performance runtime libraries and drivers for AI accelerators such as Inferentia and Trainium. You will design, develop, and deploy Neuron Runtime and other Neuron components, and contribute to profiling to optimize AI workloads across devices.

You will work with cross‑functional teams to enable performance insights, own end‑to‑end services, and participate in post‑incident reviews while

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Neuron Runtime Engineer: ML Inference & Profiling
Neuron Runtime Engineer: ML Inference & Profiling

Amazon • Cupertino (CA)

On-site
USD 165,000 - 224,000
Senior ML Infra Services Engineer for Accelerators
Senior ML Infra Services Engineer for Accelerators

Amazon • Cupertino (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Neuron Runtime Engineer — AI Inference & Profiling
Neuron Runtime Engineer — AI Inference & Profiling

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 144,000 - 194,000
Senior ML Kernel Performance Engineer - AI Accelerator
Senior ML Kernel Performance Engineer - AI Accelerator

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
AI/ML Systems Engineer - High-Performance Inference
AI/ML Systems Engineer - High-Performance Inference

Amazon Inc. • Cupertino (CA)

On-site
USD 165,000 - 224,000
Health insurance
401(k) matching
Paid time off
+1
Senior AI/ML Systems Engineer, Neuron Inference
Senior AI/ML Systems Engineer, Neuron Inference

Amazon Inc. • Cupertino (CA)

On-site
USD 193,000 - 262,000
Health insurance
401(k) matching
Paid time off
+1
Neuron Runtime Software Development Engineer , Neuron Runtime (AWS)
Neuron Runtime Software Development Engineer , Neuron Runtime (AWS)

Amazon • Seattle (WA)

On-site
USD 143,700 - 194,400
Health insurance
Dental
Vision
+4
Neuron Runtime Engineer - ML Inference & Profiling
Neuron Runtime Engineer - ML Inference & Profiling

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 144,000 - 194,000
Health insurance
RSUs
401(k) matching
+1
Senior AI/ML Inference Engineer (Neuron)
Senior AI/ML Inference Engineer (Neuron)

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
Health insurance
401(k) matching
Paid time off
Senior ML Compiler Engineer – Build Next‑Gen ML on AI Accelerators
Senior ML Compiler Engineer – Build Next‑Gen ML on AI Accelerators

Amazon • Cupertino (CA)

On-site
USD 193,300 - 261,500
Health benefits
401(k) matching
Parental leave