Open-Source ML Inference Engineer for Neuron

Amazon Inc.

Seattle (WA)

On-site

USD 144,000 - 194,000

Full time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

RSUs
Health insurance
401(k) matching
Paid time off
Parental leave

Job summary

Amazon AWS Neuron team is seeking a Software Engineer to advance AI/ML inference on Neuron chips. You will contribute to serving tech for large-scale model inference, collaborating on frameworks like vLLM and SGLang to accelerate model deployment and production readiness.

Work with cross-functional teams to optimize performance, reliability, and scalability, delivering end-to-end model support across diverse workloads and ensuring high-quality engineering practices.

Qualifications

  • 3+ years of non-internship professional software development experience.

Responsibilities

  • Contribute core serving features to vLLM and SGLang, enabling continuous batching, quantization, and distributed inference on Neuron.

Skills

C++
C#
Java
Perl
Transformer architecture
LLM fundamentals
Debugging

Education

Bachelor's degree in CS/Engineering/Math

Tools

vLLM
SGLang
TensorRT
CUDA
Triton

Job description

Amazon AWS Neuron team is seeking a Software Engineer to advance AI/ML inference on Neuron chips. You will contribute to serving tech for large-scale model inference, collaborating on frameworks like vLLM and SGLang to accelerate model deployment and production readiness.

Work with cross-functional teams to optimize performance, reliability, and scalability, delivering end-to-end model support across diverse workloads and ensuring high-quality engineering practices.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior ML Inference Engineer — Open-Source LLM Serving
Senior ML Inference Engineer — Open-Source LLM Serving

Amazon • Seattle (WA), Northern (KY)

Hybrid
USD 210,000 - 320,000
Senior ML Inference Engineer – High-Performance Serving
Senior ML Inference Engineer – High-Performance Serving

Amazon Inc. • Seattle (WA)

On-site
USD 168,000 - 227,000
ML Inference Systems Engineer — High-Performance Serving
ML Inference Systems Engineer — High-Performance Serving

Annapurna Labs (U.S.) Inc. • Seattle (WA)

On-site
USD 144,000 - 194,000
RSUs
401(k) matching
Paid time off
Sr. Software Development Engineer, Inference Team - AWS Neuron
Sr. Software Development Engineer, Inference Team - AWS Neuron

Amazon • Seattle (WA), Northern (KY)

On-site
USD 210,000 - 320,000
Senior ML Systems Engineer - AI Inference on AWS Neuron
Senior ML Systems Engineer - AI Inference on AWS Neuron

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
Software Engineer-AI/ML, Inference Team - AWS Neuron
Software Engineer-AI/ML, Inference Team - AWS Neuron

Annapurna Labs (U.S.) Inc. • Seattle (WA)

On-site
USD 144,000 - 194,000
RSUs
401(k) matching
Paid time off
Senior ML Inference Engineer — High-Performance LLM Serving
Senior ML Inference Engineer — High-Performance LLM Serving

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 168,000 - 227,000
Health insurance
401(k) matching
Paid time off
+2
ML Inference Engineer - AWS Neuron & GenAI
ML Inference Engineer - AWS Neuron & GenAI

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
Senior AI/ML Inference Engineer (Neuron)
Senior AI/ML Inference Engineer (Neuron)

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
Health insurance
401(k) matching
Paid time off
Software Engineer-AI/ML, Inference Team - AWS Neuron
Software Engineer-AI/ML, Inference Team - AWS Neuron

Amazon Inc. • Seattle (WA)

On-site
USD 144,000 - 194,000
RSUs
Health insurance
401(k) matching
+2