Senior SDE: Neuron Inference for LLM Acceleration

Amazon Inc.

Seattle (WA)

On-site

USD 144,000 - 194,000

Full time

37 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Health insurance
401(k) matching
RSUs

Job summary

Amazon AWS Neuron is seeking a senior software engineer for the Machine Learning Inference Applications team in Seattle. You will develop and optimize core building blocks of LLM inference on Neuron chips, including attention, MLP, quantization, speculative decoding, and Mixture of Experts.

Responsibilities include applying the latest research in LLM optimization to extract top performance from open source and internal models, and working across teams to deliver scalable, high-performance

Qualifications

  • 3+ years of non-internship professional software development experience.
  • 2+ years of non-internship design or architecture of systems.
  • 1+ years of SDE or related experience.
  • 1+ years of designing large-scale multi-tier distributed software.
  • 1+ years of OO design experience.
  • Bachelor's degree or foreign equivalent in CS, Engineering, Math, or related field.
  • Experience programming with at least one language (C#, C++, Java, or Perl).

Responsibilities

  • Develop and optimize core building blocks of LLM Inference (Attention, MLP, Quantization).
  • Adapt latest research in LLM optimization for Neuron chips to maximize performance.
  • Collaborate across teams to deliver scalable software for machine learning inference.

Skills

Software development
System design
OO design
C/C++/Java
Distributed systems

Education

Bachelor's degree in CS or related field

Job description

Amazon AWS Neuron is seeking a senior software engineer for the Machine Learning Inference Applications team in Seattle. You will develop and optimize core building blocks of LLM inference on Neuron chips, including attention, MLP, quantization, speculative decoding, and Mixture of Experts.

Responsibilities include applying the latest research in LLM optimization to extract top performance from open source and internal models, and working across teams to deliver scalable, high-performance

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior SDE — Neuron Inference & LLM Optimization
Senior SDE — Neuron Inference & LLM Optimization

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 144,000 - 194,000
Health insurance
401(k) matching
Paid time off
+1
SDE, Neuron Inference, Neuron Inference
SDE, Neuron Inference, Neuron Inference

Amazon Inc. • Seattle (WA)

On-site
USD 144,000 - 194,000
Health insurance
401(k) matching
RSUs
Senior ML Inference Engineer — High-Performance LLM Serving
Senior ML Inference Engineer — High-Performance LLM Serving

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 168,000 - 227,000
Health insurance
401(k) matching
Paid time off
+2
Senior ML Inference Engineer – High-Performance Serving
Senior ML Inference Engineer – High-Performance Serving

Amazon Inc. • Seattle (WA)

On-site
USD 168,000 - 227,000
Senior ML Inference Engineer — Open-Source LLM Serving
Senior ML Inference Engineer — Open-Source LLM Serving

Amazon • Seattle (WA), Northern (KY)

Hybrid
USD 210,000 - 320,000
SDE, Neuron Inference, Neuron Inference
SDE, Neuron Inference, Neuron Inference

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 144,000 - 194,000
Health insurance
401(k) matching
Paid time off
+1
ML Inference Systems Engineer — High-Performance Serving
ML Inference Systems Engineer — High-Performance Serving

Annapurna Labs (U.S.) Inc. • Seattle (WA)

On-site
USD 144,000 - 194,000
RSUs
401(k) matching
Paid time off
Sr. Software Development Engineer, Inference Team - AWS Neuron
Sr. Software Development Engineer, Inference Team - AWS Neuron

Amazon • Seattle (WA), Northern (KY)

On-site
USD 210,000 - 320,000
Software Engineer-AI/ML, Inference Team - AWS Neuron
Software Engineer-AI/ML, Inference Team - AWS Neuron

Annapurna Labs (U.S.) Inc. • Seattle (WA)

On-site
USD 144,000 - 194,000
RSUs
401(k) matching
Paid time off
Senior AI/ML Software Engineer - High-Perf Inference
Senior AI/ML Software Engineer - High-Perf Inference

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 168,000 - 227,000