Senior Neuron Runtime Software Engineer for ML Accelerators

Amazon

Cupertino (CA)

On-site

USD 150,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance (medical, dental, and
Vision and prescription coverage

Job summary

Amazon is seeking a Software Development Engineer for the Neuron Runtime Team to develop and maintain high-performance runtime libraries and drivers for ML applications and AI accelerators. You will design, deploy Neuron Runtime components and advance profiling to optimize workloads on Trainium and Inferentia devices.

You will also work on improving ML kernel and framework performance, manage the runtime life cycle, and collaborate with teams to expose key performance data through the C++

Qualifications

  • 3+ years of non-internship software development experience.
  • 2+ years of design or architecture experience (design patterns, reliability, scaling) of new and existing systems.
  • Experience programming in at least one software programming language.

Responsibilities

  • Develop and maintain high-performance runtime libraries and drivers for machine learning applications and AI accelerators.
  • Design, develop, and deploy Neuron Runtime and other Neuron components.
  • Enhance profiling capabilities to help internal and external customers optimize AI workloads across Trainium and Inferentia devices.
  • Improve performance of ML kernels and ML frameworks.
  • Manage the full development life cycle of the Neuron Runtime, ensuring scalability, reliability, and usability.
  • Collaborate with cross-functional teams to ensure the C++ compiler exposes key performance information for customers to understand and optimize performance.
  • Drive innovations to support multiple frameworks (e.g., PyTorch, JAX, XLA).

Skills

Software development experience
Design/architecture
Programming languages

Education

Bachelor’s degree in computer science

Job description

Amazon is seeking a Software Development Engineer for the Neuron Runtime Team to develop and maintain high-performance runtime libraries and drivers for ML applications and AI accelerators. You will design, deploy Neuron Runtime components and advance profiling to optimize workloads on Trainium and Inferentia devices.

You will also work on improving ML kernel and framework performance, manage the runtime life cycle, and collaborate with teams to expose key performance data through the C++

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior ML Accelerator Runtime Engineer
Senior ML Accelerator Runtime Engineer

Amazon • Seattle (WA)

On-site
USD 143,700 - 194,400
Health insurance
Dental
Vision
+4
Neuron Runtime Software Engineer - Flexible Hours
Neuron Runtime Software Engineer - Flexible Hours

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 143,000 - 195,000
Neuron Runtime Software Engineer - High-Performance ML SDKs
Neuron Runtime Software Engineer - High-Performance ML SDKs

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
Health insurance
401(k) matching
Paid time off
+1
Linux Driver & Runtime Engineer for ML Accelerator SDK
Linux Driver & Runtime Engineer for ML Accelerator SDK

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 143,000 - 195,000
Senior ML Kernel Performance Architect
Senior ML Kernel Performance Architect

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
Senior AI/ML Systems Engineer - Neuron Inference
Senior AI/ML Systems Engineer - Neuron Inference

Amazon • Cupertino (CA)

On-site
USD 193,300 - 261,500
Neuron Runtime Software Development Engineer , Neuron Runtime
Neuron Runtime Software Development Engineer , Neuron Runtime

Amazon • Cupertino (CA)

On-site
USD 150,000 - 210,000
Health insurance (medical, dental, and
Vision and prescription coverage
Engineering Manager, ML Kernel Performance
Engineering Manager, ML Kernel Performance

Amazon • Cupertino (CA)

On-site
USD 212,700 - 287,700
Health insurance
401(k) matching
Parental leave
+2
AI/ML Inference Engineer for AWS Neuron
AI/ML Inference Engineer for AWS Neuron

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
ML Kernel Performance Engineer for Neuron Accelerators
ML Kernel Performance Engineer for Neuron Accelerators

Amazon • Cupertino (CA)

On-site
USD 140,000 - 210,000