ML Engineer, Sequence Models: Scale LLMs & MLOps

Amazon Inc.

New York (NY)

On-site

USD 158,000 - 214,000

Full time

2 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

RSUs
Health insurance
401(k) matching
Paid time off

Job summary

Amazon in New York is seeking a Machine Learning Engineer for the Sequence Models team to build scalable ML infrastructure, streamline the model lifecycle from research to production, and advance MLOps for petabyte-scale data.

You will own data pipelines, optimize GPU usage, collaborate with Applied Scientists, and ensure reliable, low-latency model serving in a high-volume cloud environment. A CS/engineering degree and strong ML/LLM background are required.

Qualifications

  • 3+ years of non-internship professional software development experience.
  • 1+ years of designing and developing large-scale, multi-tiered, multi-threaded, embedded or distributed software applications, tools, systems, and services using: C#, C++, Java, or Perl.
  • Bachelor's degree or foreign equivalent in Computer Science, Engineering, Mathematics, or a related field.
  • Experience with Machine Learning and Large Language Model fundamentals, including architecture, training/inference lifecycles, and optimization of model execution.
  • Experience in developing and deploying LLMs in production on GPUs, Neuron, TPU or other AI acceleration hardware.

Responsibilities

  • Build and scale ML infrastructure across data processing, distributed training, and model serving.
  • Own the data pipelines that feed model training, including ingestion of structured and unstructured inputs, schema evolution, backfills, and data quality checks across upstream sources.
  • Partner with Applied Scientists to shorten the time from experiment to production.
  • Evolve model serving and feature delivery to support continuous experimentation.
  • Establish automated, repeatable processes for large-scale data analysis, model training, validation, and deployment.
  • Own operational excellence for high-volume, low-latency production systems, including monitoring, alarming, troubleshooting, and on-call.

Skills

Software development
Distributed systems
Java/C++/C#/Perl
ML fundamentals
LLM production

Education

Bachelor's degree

Tools

CUDA/GPUs
ML pipelines
MLOps tools

Job description

Amazon in New York is seeking a Machine Learning Engineer for the Sequence Models team to build scalable ML infrastructure, streamline the model lifecycle from research to production, and advance MLOps for petabyte-scale data.

You will own data pipelines, optimize GPU usage, collaborate with Applied Scientists, and ensure reliable, low-latency model serving in a high-volume cloud environment. A CS/engineering degree and strong ML/LLM background are required.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

ML Engineer, Sequence Models & MLOps
ML Engineer, Sequence Models & MLOps

Amazon • New York (NY)

On-site
USD 140,000 - 210,000
Machine Learning Engineer, Sequence Models (project Sequoia)
Machine Learning Engineer, Sequence Models (project Sequoia)

Amazon • New York (NY)

On-site
USD 140,000 - 210,000
ML Engineer: Production-Scale ML Pipelines
ML Engineer: Production-Scale ML Pipelines

Amazon Inc. • Factoria (WA)

On-site
USD 144,000 - 194,000
Machine Learning Engineer, Sequence Models (project Sequoia)
Machine Learning Engineer, Sequence Models (project Sequoia)

Amazon Inc. • New York (NY)

On-site
USD 158,000 - 214,000
RSUs
Health insurance
401(k) matching
+1
ML Engineer: Scalable Production ML Pipelines
ML Engineer: Scalable Production ML Pipelines

Amazon • Bellevue (WA)

On-site
USD 144,000 - 194,000
Lead ML Engineer: Scale Data Pipelines for LLMs
Lead ML Engineer: Scale Data Pipelines for LLMs

Cisco Systems, Inc. • Seattle (WA)

Hybrid
USD 180,000 - 260,000
Production AI/ML Engineer - LLM & MLOps for Global DC Ops
Production AI/ML Engineer - LLM & MLOps for Global DC Ops

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 144,000 - 194,000
Health insurance
401(k) matching
Paid time off
+2
Senior ML Engineer — Cloud, MLOps & LLMs
Senior ML Engineer — Cloud, MLOps & LLMs

GCS Recruitment • Reston (VA)

On-site
USD 120,000 - 180,000
ML DevOps Architect: Cloud & Large-Scale Compute (Remote)
ML DevOps Architect: Cloud & Large-Scale Compute (Remote)

Ignite Next GmbH • Palo Alto (CA), Northern (KY)

Hybrid
USD 140,000 - 200,000
Remote work
Office visits in Palo Alto, Paris, orW
Flexible relocation options
Remote Staff MLOps Engineer — Low-Latency ML Platform
Remote Staff MLOps Engineer — Low-Latency ML Platform

Sequen AI • United States

Remote
USD 220,000 - 280,000