Senior ML Infra Services Engineer for Accelerators

Amazon

Cupertino, Northern (CA, KY)

Hybrid

USD 140,000 - 190,000

Full time

28 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Amazon's AWS Neuron Infra Services team is seeking a Software Engineer to lead the development of tools, pipelines, and automation for optimizing machine learning workloads on Inferentia/Trainium accelerators. You will drive architecture, delivery, and customer-facing improvements across EC2/EKS/Lambda-like services within AWS or comparable platforms.

You will work with developers, hardware engineers, and users to ensure compatibility with next generation AI accelerators, while building scalable

Qualifications

  • 3+ years of non-internship professional software development experience.
  • 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience.
  • Experience programming with at least one software programming language.
  • Knowledge of system performance, memory management, and parallel computing principles.
  • Experience in debugging, profiling

Responsibilities

  • Lead the design and implementation of new tools, pipelines and automation, collaborating with developers, system architects, hardware engineers and users to ensure compatibility with existing and next-gen AI accelerators.
  • Design, implement, and maintain CI/CD pipelines to automate the software release process.
  • Collaborate with development teams to integrate new software releases.
  • Infrastructure Management: Manage and automate infrastructure provisioning.
  • Ensure high availability and scalability of systems through effective infrastructure management.
  • Monitoring and Optimization: Implement monitoring solutions to track system performance and identify bottlenecks.
  • Security and Compliance: Implement security best practices in the DevOps pipeline and conduct vulnerability assessments.

Skills

Software development
Distributed systems design
Performance analysis
Debugging & profiling
Programming languages

Tools

CI/CD pipelines
Monitoring & instrumentation

Job description

Amazon's AWS Neuron Infra Services team is seeking a Software Engineer to lead the development of tools, pipelines, and automation for optimizing machine learning workloads on Inferentia/Trainium accelerators. You will drive architecture, delivery, and customer-facing improvements across EC2/EKS/Lambda-like services within AWS or comparable platforms.

You will work with developers, hardware engineers, and users to ensure compatibility with next generation AI accelerators, while building scalable

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior ML Infra Engineer: Orchestrating AI Accelerators
Senior ML Infra Engineer: Orchestrating AI Accelerators

Amazon • Seattle (WA)

Hybrid
USD 144,000 - 194,000
Health insurance
401(k) matching
Paid time off
+1
Senior ML Accelerator Runtime Engineer
Senior ML Accelerator Runtime Engineer

Amazon • Seattle (WA)

On-site
USD 143,700 - 194,400
Health insurance
Dental
Vision
+4
Senior ML Kernel Performance Engineer - AI Accelerator
Senior ML Kernel Performance Engineer - AI Accelerator

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
Senior ML Systems Engineer - AI Inference on AWS Neuron
Senior ML Systems Engineer - AI Inference on AWS Neuron

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
Applied Scientist, ML Systems for AWS Neuron
Applied Scientist, ML Systems for AWS Neuron

Amazon • Cupertino (CA)

On-site
USD 171,600 - 222,200
RSUs
Health insurance
401(k) matching
+1
Senior ML Compiler Engineer – Build Next‑Gen ML on AI Accelerators
Senior ML Compiler Engineer – Build Next‑Gen ML on AI Accelerators

Amazon • Cupertino (CA)

On-site
USD 193,000 - 262,000
Health benefits
401(k) matching
Parental leave
ML Inference Engineer - AWS Neuron & GenAI
ML Inference Engineer - AWS Neuron & GenAI

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
ML Kernel Performance Engineer for Neuron Accelerators
ML Kernel Performance Engineer for Neuron Accelerators

Amazon • Cupertino (CA)

On-site
USD 140,000 - 210,000
Senior AI/ML Software Engineer - High-Perf Inference
Senior AI/ML Software Engineer - High-Perf Inference

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 168,000 - 227,000
SDE II, ML Infra Services, Annapurna Labs
SDE II, ML Infra Services, Annapurna Labs

Amazon.com Services LLC • Seattle (WA)

On-site
USD 180,000 - 240,000