Data Engineer: Scalable ML Pipelines & Production AI

Amazon

San Francisco, Northern (CA, KY)

Hybrid

USD 152,000 - 206,000

Full time

9 days ago
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Health insurance
401(k) matching
Paid time off
Parental leave
RSUs

Job summary

Amazon’s PXTCS team seeks a Data Engineer to collaborate with economists, data scientists, and software engineers to translate ML models into scalable production systems. You will design data pipelines using native AWS services, build APIs for model serving, and ensure data quality across diverse sources.

You will work with cross-functional teams to deliver production-ready solutions, maintaining data systems for economists and scientists while supporting secure, multi-account AWS

Qualifications

  • 3+ years of data engineering experience.
  • Experience with data modeling, warehousing and ETL pipelines.
  • Experience with AWS technologies like Redshift, S3, AWS Glue, EMR, Kinesis, Lambda, and IAM roles.

Responsibilities

  • Data Pipeline Development: Design and maintain scalable data pipelines using native AWS services; build monitoring and error handling for data workflows; optimize performance, reliability, and cost efficiency.
  • Model Productionization & API Development: Develop and maintain APIs and data serving layers that productionize science models for downstream consumption; build batch and real-time inference pipelines.
  • Data Integration & Quality: Build scalable feature extraction and processing frameworks for diverse data types; develop robust data quality and validation checks; create flexible schemas supporting evolving requirements.
  • Cross-team Collaboration: Partner with economics, data science, and software engineering teams to translate analytical requirements into production-ready solutions; participate in technical design reviews and architecture discussions.
  • Analytics & Infrastructure: Maintain layered data systems used by economists and scientists; build automated reporting solutions; work across multiple interconnected AWS accounts with security best practices.

Skills

Data engineering
Python
Java
Scala
ETL pipelines
AWS
Big data
Data modeling
SQL/NoSQL

Tools

AWS Glue
EMR
Lambda
Kinesis
Redshift
S3
IAM

Job description

Amazon’s PXTCS team seeks a Data Engineer to collaborate with economists, data scientists, and software engineers to translate ML models into scalable production systems. You will design data pipelines using native AWS services, build APIs for model serving, and ensure data quality across diverse sources.

You will work with cross-functional teams to deliver production-ready solutions, maintaining data systems for economists and scientists while supporting secure, multi-account AWS

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer - ML Production & AWS Data Pipelines
Data Engineer - ML Production & AWS Data Pipelines

Relha LLC • Seattle (WA)

On-site
USD 132,000 - 179,000
Health insurance
RSUs
401(k) matching
+2
Data Engineer: AWS Pipelines & Model Production
Data Engineer: AWS Pipelines & Model Production

Amazon • Arlington (VA)

On-site
USD 132,000 - 179,000
Health insurance
401(k) matching
Paid time off
+1
Data Engineer - AWS Pipelines & Model APIs
Data Engineer - AWS Pipelines & Model APIs

Amazon • Boston (MA)

On-site
USD 132,000 - 179,000
Health insurance
401(k) matching
Paid time off
+1
Data Engineer, AWS Pipelines & Production Analytics
Data Engineer, AWS Pipelines & Production Analytics

Amazon • Bellevue (WA)

On-site
USD 132,000 - 179,000
Health insurance
401(k) matching
Paid time off
+1
Data Platform Engineer for AI & ML Training Pipelines
Data Platform Engineer for AI & ML Training Pipelines

Amazon • Palo Alto (CA)

On-site
USD 165,000 - 224,000
Health insurance
401(k) matching
Paid time off
+1
Software Engineer: ML-Driven, Scalable Systems
Software Engineer: ML-Driven, Scalable Systems

Amazon • Seattle (WA)

On-site
USD 144,000 - 194,000
Health insurance
RSU/stock options
Data Engineer: Scalable Pipelines & AI-Driven Data
Data Engineer: Scalable Pipelines & AI-Driven Data

Amazon • New York (NY)

On-site
USD 145,000 - 197,000
ML/AI Production Systems Manager
ML/AI Production Systems Manager

Socket.dev • Bellevue (WA)

On-site
USD 185,000 - 250,000
Health insurance
401(k) matching
Parental leave
+1
Data Engineer: Scalable Pipelines for AI & Analytics
Data Engineer: Scalable Pipelines for AI & Analytics

SMX • Town of Hanover (NY)

On-site
USD 103,000 - 172,000
Health insurance
Retirement
Paid leave
Data Engineer: Scalable Pipelines for AI & Analytics
Data Engineer: Scalable Pipelines for AI & Analytics

SMX • Hanover (MD)

On-site
USD 103,000 - 172,000
Health insurance
Paid leave
Retirement plan