AI/ML Engineer, Amazon Global Data Center Ops Central Insight and Analytics Team

Amazon

Seattle (WA)

On-site

USD 144,000 - 194,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health benefits
401(k) matching
Paid time off
Parental leave

Job summary

Amazon Global Data Center Ops Central Insight and Analytics Team seeks an AI/ML Engineer to bring ML/AI models from notebooks to production. You will integrate LLMs, build RAG systems grounded in operational data, and implement robust validation, latency optimizations, and monitoring across the stack.

You will write production-grade code, deploy models, and own end-to-end tests, with a focus on reliability and observability in a large-scale data platform environment.

Qualifications

  • 3+ years of professional software development experience.
  • Bachelor's degree in Computer Science, ML, or related field.
  • 2+ years deploying ML models to production environments.
  • Strong Python proficiency and ML framework experience.
  • Experience with LLM APIs and cloud ML services.
  • Experience building data pipelines for ML.

Responsibilities

  • Build and maintain LLM-powered components and RAG pipelines.
  • Implement and optimize prompt engineering workflows and tests.
  • Deploy ML models to production with monitoring and retraining triggers.
  • Own observability: traces, latency, token usage, and SLA compliance.
  • Write comprehensive unit to end‑to‑end tests for ML pipelines.

Skills

Python
ML frameworks
LLM APIs
Cloud ML services
Data pipelines
CI/CD
Testing

Education

Bachelor's in CS/ML or related field

Tools

LangChain
Bedrock Agents
LangGraph
CrewAI
SageMaker Pipelines
MLflow
Terraform
CDK

Job description

AI/ML Engineer, Amazon Global Data Center Ops Central Insight and Analytics Team

Job ID: 10513234 | Amazon Data Services, Inc.

We are looking for an AI/ML Engineer to build, deploy, and operate the ML/AI systems that power the agentic decision intelligence workflow we are building. You are the person who takes a model from a notebook to production, builds the LLM integration layer, implements RAG pipelines, creates evaluation frameworks, and ensures our AI systems are reliable, observable, and continuously improving.

This is a hands‑on engineering role with deep ML/AI focus — you write production code that runs AI systems, not research papers. If you love the intersection of ML infrastructure, LLM applications, and production engineering, this role is for you.

Key job responsibilities
  • Build and maintain LLM-powered components: structured reasoning chains, narrative generation, recommendation rationale
  • Implement and optimize prompt engineering pipelines with version control, A/B testing, and regression detection
  • Build RAG (Retrieval-Augmented Generation) systems that ground LLM outputs in operational data, historical playbooks, and domain knowledge
  • Build guardrails, validation layers, and output parsing for LLM responses. Optimize latency, cost, and quality trade-offs across LLM providers
  • Deploy ML models to production. Implement model monitoring: drift detection, performance degradation alerts, automated retraining triggers
  • Build A/B testing infrastructure for model experiments. Manage model versioning, rollback, and canary deployment. Ensure SLA compliance for inference latency and availability
  • Own the operational health of AI/ML services: monitoring, alarming, on‑call, incident response, observability across the AI stack (prompt traces, latency histograms, token usage, error rates)
  • Write comprehensive tests (unit, integration, end‑to‑end) for ML pipelines
Basic Qualifications
  • 3+ years of non‑internship professional software development experience
  • Bachelor's degree in Computer Science, Machine Learning, or related field (or equivalent experience)
  • 2+ years deploying ML models to production environments
  • Strong Python proficiency + experience with ML frameworks
  • Experience with LLM APIs and prompt engineering
  • Experience with cloud ML services
  • Experience building data pipelines for ML (feature engineering, preprocessing, training data management)
  • Solid software engineering fundamentals (testing, CI/CD, code review, production operations)
Preferred Qualifications
  • 3+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience
  • Experience building RAG systems (vector databases, embedding models, retrieval pipelines)
  • Experience with agent/orchestration frameworks (LangChain, LangGraph, CrewAI, Bedrock Agents, or custom)
  • Experience with ML evaluation frameworks (especially for generative AI / LLM outputs)
  • Experience with time-series ML (forecasting, anomaly detection)
  • Experience with MLOps tooling (MLflow, SageMaker Pipelines, Step Functions, feature stores)
  • Experience with infrastructure-as-code (CDK, CloudFormation, Terraform)
  • Background in operational/infrastructure environments

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

The base salary range for this position is listed below. Your Amazon package will include sign‑on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave.

Learn more about our benefits at https://amazon.jobs/en/benefits.

USA, WA, Seattle - 143,700.00 - 194,400.00 USD annually

Important FAQs for current Government employees

Before proceeding, please review the following FAQs

https://www.amazon.jobs/en/faqs#faqs-for-us-government-employees

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI/ML Engineer, Amazon Global Data Center Ops Central Insight and Analytics Team
AI/ML Engineer, Amazon Global Data Center Ops Central Insight and Analytics Team

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 144,000 - 194,000
Health insurance
401(k) matching
Paid time off
+2
Senior Applied Scientist, Amazon Global Data Center Ops Central Insight and Analytics Team
Senior Applied Scientist, Amazon Global Data Center Ops Central Insight and Analytics Team

Amazon • Seattle (WA)

On-site
USD 167,000 - 226,000
Software Development Engineer, AAIS
Software Development Engineer, AAIS

Amazon • Seattle (WA)

On-site
USD 143,700 - 194,400
Health insurance
401(k) matching
Paid time off
+1
Delivery Consultant - AI/ML, AWS Professional Services - HCLS
Delivery Consultant - AI/ML, AWS Professional Services - HCLS

Amazon • Dallas (TX)

On-site
USD 150,000 - 210,000
Applied Scientist, Applied AI Solutions (AWS)
Applied Scientist, Applied AI Solutions (AWS)

Amazon • Seattle (WA)

On-site
USD 143,000 - 193,000
Health insurance
401(k) matching
Paid time off
Sr. GenAI/ML Specialist Solutions Architect, AGS Specialist Solutions Architects (AWS)
Sr. GenAI/ML Specialist Solutions Architect, AGS Specialist Solutions Architects (AWS)

Amazon • San Francisco (CA)

On-site
USD 177,000 - 239,000
Health insurance
RSUs
401(k) matching
+1
AI/ML Engineer, Gateway
AI/ML Engineer, Gateway

Amazon • Seattle (WA)

On-site
USD 144,000 - 194,000
Health insurance
401(k) matching
Paid time off
+2
Senior Applied Scientist, Amazon Global Data Center Ops Central Insight and Analytics Team
Senior Applied Scientist, Amazon Global Data Center Ops Central Insight and Analytics Team

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 167,000 - 226,000
Health insurance
RSUs
Sign-on payments
+3
Delivery Consultant- AI/ML, Data & Machine Learning (DML)
Delivery Consultant- AI/ML, Data & Machine Learning (DML)

Amazon • Arlington (VA)

On-site
USD 131,300 - 177,600
Health insurance
401(k) matching
Paid time off
+1
Sr. Delivery Consultant - AI/ML, AWS Professional Services
Sr. Delivery Consultant - AI/ML, AWS Professional Services

Amazon Web Services (AWS) • Arlington (VA)

On-site
USD 154,000 - 208,000