Neuron Runtime Engineer - ML Inference & Profiling

Amazon Web Services (AWS)

Seattle (WA)

On-site

USD 144,000 - 194,000

Full time

6 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Health insurance
RSUs
401(k) matching
Paid time off

Job summary

Annapurna Labs (U.S.) Inc. in Seattle, WA is seeking a Software Development Engineer for the Neuron Runtime Team.

You will build and maintain high‑performance runtime libraries and drivers for ML accelerators, and collaborate across teams to ensure the C++ compiler exposes key performance insights to customers. The role focuses on scalable distributed systems, profiling performance across Trainium and Inferentia devices, and integrating support for PyTorch, JAX, and XLA while owning deployment

Qualifications

  • 3+ years of non‑internship professional software development experience.
  • 2+ years of non‑internship design or architecture experience (design patterns, reliability, scaling).
  • Experience programming with at least one software programming language.

Responsibilities

  • Develop and maintain high‑performance runtime libraries and drivers for ML accelerators.
  • Design, develop, and deploy Neuron Runtime and related components.
  • Improve profiler to support multiple frameworks and optimize ML workloads.

Skills

Software development
System design
Programming language

Education

Bachelor's degree in computer science

Job description

Annapurna Labs (U.S.) Inc. in Seattle, WA is seeking a Software Development Engineer for the Neuron Runtime Team.

You will build and maintain high‑performance runtime libraries and drivers for ML accelerators, and collaborate across teams to ensure the C++ compiler exposes key performance insights to customers. The role focuses on scalable distributed systems, profiling performance across Trainium and Inferentia devices, and integrating support for PyTorch, JAX, and XLA while owning deployment

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Neuron Runtime Engineer — AI Inference & Profiling
Neuron Runtime Engineer — AI Inference & Profiling

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 144,000 - 194,000
Neuron Runtime Engineer: ML Inference & Profiling
Neuron Runtime Engineer: ML Inference & Profiling

Amazon • Cupertino (CA)

On-site
USD 165,000 - 224,000
Neuron Runtime Software Engineer — Flexible Hours
Neuron Runtime Software Engineer — Flexible Hours

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
Health insurance
401(k) matching
Paid time off
+2
Senior ML Compiler Engineer – Neuron
Senior ML Compiler Engineer – Neuron

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 168,000 - 227,000
Senior AI/ML Inference Engineer (Neuron)
Senior AI/ML Inference Engineer (Neuron)

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
Health insurance
401(k) matching
Paid time off
Senior ML Accelerator Runtime Engineer
Senior ML Accelerator Runtime Engineer

Amazon • Seattle (WA)

On-site
USD 143,700 - 194,400
Health insurance
Dental
Vision
+4
Senior ML Systems Engineer - AI Inference on AWS Neuron
Senior ML Systems Engineer - AI Inference on AWS Neuron

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
ML & Compiler Engineer for AI Accelerators
ML & Compiler Engineer for AI Accelerators

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
Health insurance
401(k) matching
RSU/stock options
+1
ML Inference Engineer - AWS Neuron & GenAI
ML Inference Engineer - AWS Neuron & GenAI

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
Senior AI/ML Software Engineer - High-Perf Inference
Senior AI/ML Software Engineer - High-Perf Inference

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 168,000 - 227,000