Data Engineer: Production ML Pipelines on AWS

Amazon

Seattle (WA)

On-site

USD 132,000 - 179,000

Full time

28 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Amazon.com Services LLC is seeking an experienced Data Engineer to join the PXTCS team, working with economists, data scientists, software engineers, and applied scientists to turn ML/Generative AI models into reliable production systems.

You will enhance data architecture, standardize metrics, and deliver end-to-end data engineering solutions for complex analytical problems, collaborating across teams to translate requirements into scalable insights.

Qualifications

  • Knowledge of professional software engineering & best practices for full software development life cycle, including coding standards, architectures, code reviews, source control, CI/CD, testing and operations.
  • 3+ years of data engineering experience.
  • Experience in at least one modern scripting or programming language such as Python, Java, Scala, or NodeJS.
  • Experience with data modeling, warehousing and building ETL pipelines.
  • Experience with AWS technologies like Redshift, S3, Glue, EMR, Kinesis, FireHose, Lambda, and IAM roles and permissions.
  • Experience with non-relational databases / data stores (object storage, document or key-value stores, graph databases, column-family databases).

Responsibilities

  • Data Pipeline Development: Design and maintain scalable data pipelines using native AWS services (Glue, EMR, Lambda); build monitoring and error handling for data workflows; optimize performance, reliability, and cost efficiency.
  • Model Productionization & API Development: Develop and maintain APIs and data serving layers that productionize science models for downstream consumption; build batch and real-time inference pipelines.
  • Data Integration & Quality: Build scalable feature extraction and processing frameworks for diverse data types; develop robust data quality and validation checks; create flexible schemas supporting evolving requirements.
  • Cross-team Collaboration: Partner with economics, data science, and software engineering teams to translate analytical requirements into production-ready solutions; participate in technical design reviews and architecture discussions.
  • Analytics & Infrastructure: Maintain layered data systems used by economists and scientists; build automated reporting solutions; work across multiple interconnected AWS accounts with security best practices.

Skills

Data engineering
Software engineering practices
Python/Java/Scala/NodeJS

Tools

AWS services (Redshift, S3, Glue, EMR, Kinesis, Lambda)
Non-relational databases

Job description

Amazon.com Services LLC is seeking an experienced Data Engineer to join the PXTCS team, working with economists, data scientists, software engineers, and applied scientists to turn ML/Generative AI models into reliable production systems.

You will enhance data architecture, standardize metrics, and deliver end-to-end data engineering solutions for complex analytical problems, collaborating across teams to translate requirements into scalable insights.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer - ML Production & AWS Data Pipelines
Data Engineer - ML Production & AWS Data Pipelines

Relha LLC • Seattle (WA)

On-site
USD 132,000 - 179,000
Health insurance
RSUs
401(k) matching
+2
Data Engineer: Scalable ML Pipelines & Production AI
Data Engineer: Scalable ML Pipelines & Production AI

Amazon • San Francisco (CA), Northern (KY)

Hybrid
USD 152,000 - 206,000
Health insurance
401(k) matching
Paid time off
+2
Data Engineer - Production ML Pipelines for People Ops
Data Engineer - Production ML Pipelines for People Ops

Amazon • Arlington (VA)

On-site
USD 132,000 - 179,000
Health insurance
401(k) matching
Paid time off
+2
Data Engineer II: Manufacturing Analytics & ML Pipelines
Data Engineer II: Manufacturing Analytics & ML Pipelines

Amazon • Bellevue (WA)

On-site
USD 132,000 - 179,000
Health insurance
PTO
401(k)
Data Platform Engineer for AI & ML Training Pipelines
Data Platform Engineer for AI & ML Training Pipelines

Amazon • Palo Alto (CA)

On-site
USD 165,000 - 224,000
Health insurance
401(k) matching
Paid time off
+1
Data Engineer: Scalable ETL & AWS Data Pipelines
Data Engineer: Scalable ETL & AWS Data Pipelines

Amazon • Seattle (WA)

On-site
USD 132,000 - 179,000
Data Engineer — AI-Powered Data Platform
Data Engineer — AI-Powered Data Platform

Amazon Web Services (AWS) • Arlington (VA)

On-site
USD 132,000 - 179,000
Data Engineer: Scalable Pipelines for AI & Analytics
Data Engineer: Scalable Pipelines for AI & Analytics

SMX • Town of Hanover (NY)

On-site
USD 103,000 - 172,000
Health insurance
Retirement
Paid leave
Data Engineer: Scalable Pipelines for AI & Analytics
Data Engineer: Scalable Pipelines for AI & Analytics

SMX • Hanover (MD)

On-site
USD 103,000 - 172,000
Health insurance
Paid leave
Retirement plan
ML Engineer: Cloud ML Pipelines & Production
ML Engineer: Cloud ML Pipelines & Production

Tata Consultancy Services • Conroe (TX)

On-site
USD 110,000 - 140,000
Discretionary Incentive
Medical Coverage
Parental Leaves
+4