SDE, AI/ML Networking — Disaggregated Inference Expert

Amazon

Cupertino (CA)

On-site

USD 165,000 - 224,000

Full time

8 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Health insurance
RSUs
Paid time off

Job summary

Annapurna Labs, an integral part of AWS, is seeking a Software Development Engineer to push the frontier of AI/ML networking with disaggregated inference. You will build and optimize data-movement software across accelerators and servers, profiling workloads to approach hardware limits and collaborating across stacks from transport to inference frameworks.

You will join a team that designs hardware and software for EC2 infrastructure, focusing on ML/HPC workloads on AWS, with mentorship and

Qualifications

  • 3+ years of non‑internship professional software development experience.
  • 2+ years of non‑internship design or architecture experience.
  • Experience programming with at least one software language (C/C++ highlighted).
  • Experience with C/C++.

Responsibilities

  • Build and optimize the low-level data-movement software that transfers KV cache and activations across accelerators, servers, and heterogeneous memory — over AWS's highest-performance network fabric.
  • Profile real workloads, find the true bottleneck, and close the gap between "it works" and "it runs fast" — pushing components toward the hardware's limit.
  • Work across the stack — from network transport up to the inference frameworks — learning from the teams building the chips, runtime, and models.
  • Deliver features that ship to our largest clusters, for our largest customers, serving the largest AI models in production.

Skills

C/C++
Software development
System design/architecture
Performance optimization

Education

Bachelor's degree in computer science or equivalent

Tools

RDMA/InfiniBand
libfabric/UCX
NCCL
MPI

Job description

Annapurna Labs, an integral part of AWS, is seeking a Software Development Engineer to push the frontier of AI/ML networking with disaggregated inference. You will build and optimize data-movement software across accelerators and servers, profiling workloads to approach hardware limits and collaborating across stacks from transport to inference frameworks.

You will join a team that designs hardware and software for EC2 infrastructure, focusing on ML/HPC workloads on AWS, with mentorship and

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer - AI/ML Disaggregated Inference
Software Engineer - AI/ML Disaggregated Inference

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
Senior AI/ML Systems Engineer - Disaggregated Inference
Senior AI/ML Systems Engineer - Disaggregated Inference

Energy Jobline ZR • Cupertino (CA)

On-site
USD 193,000 - 262,000
Health insurance
401(k) matching
Paid time off
+1
Senior Systems Engineer, AI Inference & HPC Networking
Senior Systems Engineer, AI Inference & HPC Networking

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
RSUs
Health insurance
401(k) matching
+1
Systems Engineer, High‑Speed AI Inference Networking
Systems Engineer, High‑Speed AI Inference Networking

Amazon • Cupertino (CA)

On-site
USD 165,000 - 224,000
Health insurance
RSUs / stock options
401(k) matching
+2
SDE I, ML Infra & AI Accelerators
SDE I, ML Infra & AI Accelerators

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 127,000 - 185,000
Senior AI/ML Systems Engineer - Flexible Hours & Networking
Senior AI/ML Systems Engineer - Flexible Hours & Networking

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
Senior ML Network Stack Engineering Manager
Senior ML Network Stack Engineering Manager

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 220,000 - 298,000
Health insurance
401(k) matching
Paid time off
+1
Senior AI/ML Software Engineer - High-Perf Inference
Senior AI/ML Software Engineer - High-Perf Inference

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 168,000 - 227,000
AI/ML Network Infrastructure Engineer I
AI/ML Network Infrastructure Engineer I

Amazon • Cupertino (CA)

On-site
USD 127,000 - 185,000
Lead ML Network Stack Engineer for Scalable EC2 AI
Lead ML Network Stack Engineer for Scalable EC2 AI

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
Health insurance
RSUs and sign-on options
401(k) matching
+2