Linux Driver & Runtime Engineer for ML Accelerator SDK

Amazon Web Services (AWS)

Austin (TX)

On-site

USD 143,700 - 194,400

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Annapurna Labs (U.S.) Inc. is seeking a Software Engineer for the AWS Neuron SDK team in Austin, TX. You will build the runtime stack and device drivers to maximize performance on AWS Inferentia and Trainium accelerators.

The role emphasizes Linux driver development, low-level software, pre-silicon validation, and interfaces with hardware teams to push throughput, reliability, and cost efficiencies across inference and training workloads.

Qualifications

  • 3+ years of professional software development experience.
  • 2+ years of design or architecture experience for new or existing systems.
  • Experience programming in at least one software language.

Responsibilities

  • Define interfaces and develop runtime stack and driver for accelerators.
  • Work in pre-silicon environments (emulation) to shift-left readiness and validation content.
  • Collaborate with hardware and software teams to maximize throughput and cost efficiency.
  • Deliver high-quality, scalable software in a fast-paced environment.

Skills

Software development
System design
Programming languages

Education

Bachelor's degree in Computer Science or equivalent

Tools

Low level drivers

Job description

Annapurna Labs (U.S.) Inc. is seeking a Software Engineer for the AWS Neuron SDK team in Austin, TX. You will build the runtime stack and device drivers to maximize performance on AWS Inferentia and Trainium accelerators.

The role emphasizes Linux driver development, low-level software, pre-silicon validation, and interfaces with hardware teams to push throughput, reliability, and cost efficiencies across inference and training workloads.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Neuron Runtime Software Engineer for ML Accelerators
Senior Neuron Runtime Software Engineer for ML Accelerators

Amazon • Cupertino (CA)

On-site
USD 150,000 - 210,000
Health insurance (medical, dental, and
Vision and prescription coverage
Senior ML Accelerator Runtime Engineer
Senior ML Accelerator Runtime Engineer

Amazon • Seattle (WA)

On-site
USD 143,700 - 194,400
Health insurance
Dental
Vision
+4
ML Kernel Performance Engineer for AI Accelerators
ML Kernel Performance Engineer for AI Accelerators

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
Health insurance
RSUs
401(k) matching
+1
Neuron Runtime Software Engineer - Flexible Hours
Neuron Runtime Software Engineer - Flexible Hours

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 143,000 - 195,000
AI/ML Inference Engineer for AWS Neuron
AI/ML Inference Engineer for AWS Neuron

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
Runtime/Driver Software Development Engineer, Neuron Runtime
Runtime/Driver Software Development Engineer, Neuron Runtime

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 143,000 - 195,000
ML Kernel Performance Engineering Manager
ML Kernel Performance Engineering Manager

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 212,000 - 288,000
Neuron Runtime Software Engineer - High-Performance ML SDKs
Neuron Runtime Software Engineer - High-Performance ML SDKs

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
Health insurance
401(k) matching
Paid time off
+1
AI/ML SDE — High-Performance Inference on Custom Hardware
AI/ML SDE — High-Performance Inference on Custom Hardware

Amazon • Cupertino (CA)

On-site
USD 180,000 - 240,000
Senior ML Kernel Optimization Engineer
Senior ML Kernel Optimization Engineer

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000