Machine Learning - Compiler Engineer , AWS Neuron, Annapurna Labs

Amazon Web Services (AWS)

Cupertino (CA)

On-site

USD 165,200 - 223,600

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
RSUs
Sign-on bonus

Job summary

AWS is seeking a software engineer in the Neuron Compiler team to build the next-generation compiler translating PyTorch, TensorFlow, and JAX models for AWS Inferentia and Trainium servers in the cloud.

You will optimize compiler performance, collaborate with chip architects and ML apps teams, and contribute to open-source projects to advance ML workloads on AWS hardware.

Qualifications

  • 3+ years of non-internship software development experience.
  • 2+ years of non-internship design or architecture experience of new and existing systems.
  • Experience programming with at least one software programming language.

Responsibilities

  • Design, implement, test, deploy, and maintain software solutions that improve the Neuron compiler’s performance, stability, and UI.
  • Collaborate with chip architects, runtime/OS engineers, scientists, and ML apps teams to deploy ML models on AWS accelerators.
  • Engage with open-source projects like StableHLO, OpenXLA, and MLIR to advance ML workload optimization on AWS hardware.

Skills

Software development
System design
C++/Java

Education

Master's degree or PhD in Computer Science

Tools

LLVM
MLIR
PyTorch
JAX
Bazel
CMake

Job description

Do you want to be part of the AI revolution? At AWS our vision is to make deep learning pervasive for everyday developers and to democratize access to AI hardware and software infrastructure. To deliver on that vision, we’ve created innovative software and hardware solutions that make it possible. AWS Neuron is the SDK that optimizes the performance of complex ML models executed on AWS Inferentia and Trainium, our custom chips designed to accelerate deep‑learning workloads.

This role is for a software engineer in the Compiler team for AWS Neuron. You will build next‑generation Neuron compiler that transforms ML models written in frameworks such as PyTorch, TensorFlow, and JAX to run on AWS Inferentia and Trainium based servers in the Amazon cloud. Your work will involve solving hard compiler optimization problems to achieve optimum performance for a variety of ML model families, including massive‑scale large language models like Llama and Deepseek, as well as stable diffusion, vision transformers, and multi‑model setups. You will deeply understand how these models work internally to inform compiler design decisions, communicate with internal and external stakeholders, and participate in pre‑silicon design to bring new products and features to market. The goal is to make the Neuron compiler highly performant and easy‑to‑use.

Key responsibilities include designing, implementing, testing, deploying, and maintaining innovative software solutions that improve the Neuron compiler’s performance, stability, and user interface. You will work closely with chip architects, runtime/OS engineers, scientists, and ML apps teams to deploy state‑of‑the‑art ML models on AWS accelerators with optimal cost/performance benefits. You will also engage with open‑source projects such as StableHLO, OpenXLA, and MLIR to pioneer advanced ML workload optimization on AWS hardware, build features that deliver great developer experiences, and create tools to analyze numerical errors and resolve compiler defects.

Experience in object‑oriented languages like C++/Java is a must. Experience with compilers or building ML models on accelerators (e.g., GPUs) is preferred but not required. Familiarity with OpenXLA, StableHLO, and MLIR is a bonus.

Basic Qualifications
  • 3+ years of non‑internship professional software development experience
  • 2+ years of non‑internship design or architecture experience of new and existing systems (design patterns, reliability, scaling)
  • Experience programming with at least one software programming language
Preferred Qualifications
  • Master's degree or PhD in Computer Science or a related technical field
  • 3+ years of experience writing production‑grade code in object‑oriented languages such as C++ or Java
  • Experience in compiler design for CPU/GPU/Vector engines or ML‑accelerators
  • Experience with open‑source compiler toolsets like LLVM or MLIR
  • Experience with technologies such as PyTorch, OpenXLA, StableHLO, JAX, TVM, deep‑learning models, and algorithms
  • Experience with modern build systems such as Bazel or CMake

Amazon is an equal‑opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Benefits: The base salary range for this position is USD 165,200.00 – 223,600.00 annually. Your Amazon package will include sign‑on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D, optional Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more at https://amazon.jobs/en/benefits.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr. Machine Learning - Compiler Engineer Iii, Aws Neuron, Annapurna Labs
Sr. Machine Learning - Compiler Engineer Iii, Aws Neuron, Annapurna Labs

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
Health benefits
401(k) matching
Parental leave
Sr. Machine Learning - Compiler Engineer III, AWS Neuron, Annapurna Labs
Sr. Machine Learning - Compiler Engineer III, AWS Neuron, Annapurna Labs

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 168,000 - 227,000
ML Compiler Engineer, Annapurna Labs
ML Compiler Engineer, Annapurna Labs

Amazon • Cupertino (CA)

On-site
USD 165,000 - 224,000
Software Development Manager - Compiler
Software Development Manager - Compiler

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 213,000 - 288,000
Health insurance
RSUs
Sign-on payments
Software Development Manager, ML Accelerators, AWS Neuron, Annapurna Labs
Software Development Manager, ML Accelerators, AWS Neuron, Annapurna Labs

Amazon • Seattle (WA)

On-site
USD 166,400 - 287,700
Flexible working hours
Mentorship programs
Inclusive culture
Software Development Manager - Compiler (AWS)
Software Development Manager - Compiler (AWS)

Amazon • Cupertino (CA)

On-site
USD 180,000 - 260,000
Software Development Manager, ML Accelerators, AWS Neuron, Annapurna Labs
Software Development Manager, ML Accelerators, AWS Neuron, Annapurna Labs

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 184,000 - 251,000
Software Development Manager , AWS Neuron, Annapurna Labs
Software Development Manager , AWS Neuron, Annapurna Labs

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 150,000 - 200,000
Applied Scientist, Neuron ARG, Annapurna ML
Applied Scientist, Neuron ARG, Annapurna ML

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 143,000 - 193,000
Health insurance
RSUs
401(k) matching
+2
Applied Scientist — ML Systems & Compiler for Neuron
Applied Scientist — ML Systems & Compiler for Neuron

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 133,000 - 222,000