Senior ML Inference Engineer for Trainium & Neuron

Artha Nexgen

Cupertino (CA)

Hybrid

USD 193,000 - 262,000

Full time

6 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Amazon’s AWS Neuron team seeks a Sr. Software Development Engineer for Ai/ML inference model enablement on Trainium accelerators. You will onboard and optimize LLMs for inference on Trainium, improve enabling speed, and advance usability with tools and automation.

You will lead design reviews, debug performance across the stack, mentor engineers, and drive architectural decisions shaping Neuron’s inference capabilities, in a collaborative environment with applied scientists and product managers.

Qualifications

  • 5+ years of professional software development experience.
  • 5+ years leading design or architecture of systems.
  • 5+ years full software development life cycle experience.
  • 5+ years of non‑internship professional software development.
  • Experience as a mentor, tech lead or leading an engineering team.
  • Master’s degree in CS or related field preferred.

Responsibilities

  • Deliver high-performance models using distributed inference libraries.
  • Drive architectural decisions that impact the Neuron serving stack.
  • Mentor team members and provide technical leadership across multiple work streams.
  • Collaborate with customers, product owners, and engineering teams to define technical strategy.
  • Author technical documentation, design proposals, and architectural guidelines.

Skills

Java
C++
C#
System architecture
Mentorship

Education

Master's degree in Computer Science

Job description

Amazon’s AWS Neuron team seeks a Sr. Software Development Engineer for Ai/ML inference model enablement on Trainium accelerators. You will onboard and optimize LLMs for inference on Trainium, improve enabling speed, and advance usability with tools and automation.

You will lead design reviews, debug performance across the stack, mentor engineers, and drive architectural decisions shaping Neuron’s inference capabilities, in a collaborative environment with applied scientists and product managers.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI SDK Lead for Neuron on Trainium/Inferentia
Senior AI SDK Lead for Neuron on Trainium/Inferentia

Relha LLC • Seattle (WA), Northern (KY)

Hybrid
USD 220,000 - 298,000
Senior ML Systems Engineer - AI Inference on AWS Neuron
Senior ML Systems Engineer - AI Inference on AWS Neuron

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
Senior AI Software Manager: Neuron SDK Lead
Senior AI Software Manager: Neuron SDK Lead

Socket.dev • Seattle (WA)

On-site
USD 220,000 - 342,000
Health insurance
RSUs (Restricted Stock Units)
401(k) matching
+1
Applied Scientist, ML Systems for AWS Neuron
Applied Scientist, ML Systems for AWS Neuron

Amazon • Cupertino (CA)

On-site
USD 171,600 - 222,200
RSUs
Health insurance
401(k) matching
+1
AI/ML Systems Engineer - Neuron Inference
AI/ML Systems Engineer - Neuron Inference

Amazon • Seattle (WA)

On-site
USD 144,000 - 194,000
Senior AI/ML Software Engineer - Neuron Optimizations
Senior AI/ML Software Engineer - Neuron Optimizations

Annapurna Labs (U.S.) Inc. • Seattle (WA)

On-site
USD 180,000 - 230,000
ML Inference Engineer - AWS Neuron & GenAI
ML Inference Engineer - AWS Neuron & GenAI

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
AI/ML Software Engineer for High-Performance Training
AI/ML Software Engineer for High-Performance Training

Amazon • Cupertino (CA)

On-site
USD 165,000 - 224,000
Senior ML Accelerator Runtime Engineer
Senior ML Accelerator Runtime Engineer

Amazon • Seattle (WA)

On-site
USD 143,700 - 194,400
Health insurance
Dental
Vision
+4
Machine Learning
Machine Learning

Artha Nexgen • Cupertino (CA)

Hybrid
USD 193,000 - 262,000