Senior Data Engineer, Applied AI Solutions

Amazon Web Services (AWS)

Seattle (WA)

On-site

USD 155,000 - 209,000

Full time

17 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Health insurance
401(k) matching
Paid time off
Parental leave

Job summary

Amazon Development Center U.S., Inc. in Seattle is seeking a Senior Data Engineer to design and maintain a next-generation data infrastructure that serves both human analysts and AI systems.

You will bridge traditional data warehousing with AI technologies, ensuring data accuracy, lineage, and observability at scale, while collaborating across data scientists, engineers, and business stakeholders.

Qualifications

  • 7+ years of data engineering experience.
  • Experience building data infrastructure for AI systems and autonomous agents.
  • Experience with GenAI data patterns end to end: chunking, embeddings, vector stores.
  • Experience maintaining datasets and feature pipelines for ML/GenAI training and inference.
  • Experience with data quality, lineage, and observability at scale.
  • Strong Python (or Scala/Java) and advanced SQL skills.

Responsibilities

  • Design and operate production data pipelines and warehouses.
  • Build data infrastructure that serves AI systems and autonomous agents.
  • Develop end-to-end GenAI data patterns: chunking, embeddings, vector stores.
  • Create datasets/feature pipelines for ML/GenAI training and inference.
  • Implement data quality, lineage, and observability for AI workloads.
  • Collaborate with data scientists, engineers, and stakeholders.

Skills

Python
SQL
Data engineering
Mentoring
ETL/ELT
GenAI data patterns
Batch/Streaming ETL
Cloud data pipelines

Education

Bachelor's degree in CS/Engineering/Analytics/Math/Statistics/IT

Tools

Amazon Redshift
Glue
EMR/Spark
SageMaker
Bedrock
Airflow
Step Functions
Glue Workflows
S3

Job description

Description

The newest business group in AWS, Applied AI Solutions are built by AWS and AWS Partners to deliver applied AI solutions that leverage Amazon’s operational expertise and that businesses love and trust for their day-to-day success. Our ambition is to become a partner which companies can rely on to run their business every day, putting AI to work delivering better customer experience, operational excellence and speed. We are seeking a Senior Data Engineer to design, build and maintain our next-generation data infrastructure - one that seamlessly serves both human analysts and AI systems. This role sits at the intersection of traditional enterprise data warehousing and innovative AI technologies, requiring someone who can bridge these worlds to create a unified, future-proof data ecosystem. As a key member of our data team, you'll collaborate across organizational boundaries with data scientists, engineers, analytics teams, and business stakeholders to develop innovative and scalable solutions that push the boundaries of what's possible with our data assets. You'll be responsible for ensuring our datasets maintain the highest levels of accuracy, consistency, and observability - implementing comprehensive monitoring, lineage tracking, and self-healing mechanisms that maintain data quality at scale. Your infrastructure will support both analysts / scientists and autonomous AI agents with equal effectiveness, requiring thoughtful interfaces, documentation, and metadata that serve both audiences. In this role, you'll champion a forward-thinking approach to data infrastructure that anticipates the evolving needs of AI systems while maintaining the reliability and performance that business operations demand. You'll help shape our technical roadmap for data systems that will serve as the foundation for our organization's AI transformation journey.

Key job responsibilities
  • 5+ years of data engineering, building and operating production pipelines and warehouses.
  • Experience building data infrastructure that serves AI systems and autonomous agents, not just human analysts, including machine-consumable interfaces, metadata, and documentation.
  • Experience with GenAI data patterns end to end: chunking, embeddings, and vector stores for retrieval-augmented generation.
  • Experience building and maintaining datasets and feature pipelines for ML/GenAI training, fine-tuning, and inference (Amazon SageMaker, Bedrock, or equivalent).
  • Experience implementing data quality, lineage, and observability that AI workloads depend on including validation, freshness/anomaly monitoring, and alerting at scale.
  • 5+ years of Python (or Scala/Java) and advanced SQL, including performance tuning at scale.
  • Experience with batch and streaming ETL/ELT on AWS (Glue, EMR/Spark, S3, Athena) and a production cloud data warehouse (Amazon Redshift or equivalent).
  • Experience designing data models and schemas for analytical, operational, and AI/retrieval workloads.
  • Experience with workflow orchestration (Step Functions, Airflow, or Glue Workflows).
Basic Qualifications
  • 7+ years of data engineering experience
  • Experience with data modeling, warehousing and building ETL pipelines
  • Experience with SQL
  • Experience in at least one modern scripting or programming language, such as Python, Java, Scala, or NodeJS
  • Experience mentoring team members on best practices
  • Experience with MPP databases such as Amazon Redshift
  • Experience building/operating highly available, distributed systems of data extraction, ingestion, and processing of large data sets
  • Experience building data infrastructure that serves AI systems and autonomous agents, not just human analysts, including machine-consumable interfaces, metadata, and documentation.
  • Experience with GenAI data patterns end to end: chunking, embeddings, and vector stores for retrieval-augmented generation.
  • Experience building and maintaining datasets and feature pipelines for ML/GenAI training, fine-tuning, and inference (Amazon SageMaker, Bedrock, or equivalent).
  • Experience implementing data quality, lineage, and observability that AI workloads depend on including validation, freshness/anomaly monitoring, and alerting at scale.
Preferred Qualifications
  • Experience with big data technologies such as: Hadoop, Hive, Spark, EMR
  • Experience operating large data warehouses
  • Experience providing technical leadership and mentoring other engineers for best practices on data engineering
  • Bachelor's degree in computer science, engineering, analytics, mathematics, statistics, IT or equivalent
  • Knowledge of distributed systems as it pertains to data storage and computing

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https:\/\/amazon.jobs\/content\/en\/how-we-hire\/accommodations for more information. If the country\/region you’re applying in isn’t listed, please contact your Recruiting Partner.

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https:\/\/amazon.jobs\/en\/benefits.

USA, WA, Seattle - 154,600.00 - 209,100.00 USD annually

Company - Amazon Development Center U.S., Inc.

Job ID: A10559413

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Data Engineer, Applied AI Solutions
Senior Data Engineer, Applied AI Solutions

Amazon Inc. • Seattle (WA)

On-site
USD 155,000 - 209,000
Data Engineer, WW Ops Finance - S&A
Data Engineer, WW Ops Finance - S&A

Amazon Inc. • Omaha (NE), Northern (KY)

Hybrid
USD 132,000 - 179,000
Software Development Engineer, AWS Analytics Engineering
Software Development Engineer, AWS Analytics Engineering

Amazon • Seattle (WA)

On-site
USD 144,000 - 194,000
Data Engineer, WW Ops Finance - S&A
Data Engineer, WW Ops Finance - S&A

Amazon • Factoria (WA)

On-site
USD 132,000 - 179,000
Health insurance
401(k) matching
Paid time off
Software Development Engineer, Eva Development Team
Software Development Engineer, Eva Development Team

Amazon Web Services (AWS) • Bellevue (WA)

On-site
USD 144,000 - 194,000
Sr BIE, ASP Business Intelligence
Sr BIE, ASP Business Intelligence

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 130,000 - 176,000
Health insurance
RSUs
401(k) matching
Senior Software Development Engineer, AWS Analytics Engineering
Senior Software Development Engineer, AWS Analytics Engineering

Amazon • Seattle (WA)

On-site
USD 168,000 - 227,000
Sr. Data Engineer, Deal Tooling and Insights, Strategic Customer Engagements
Sr. Data Engineer, Deal Tooling and Insights, Strategic Customer Engagements

Amazon Web Services (AWS) • Arlington (VA)

On-site
USD 155,000 - 209,000
Sr. Data Engineer, Strategic Customer Engagements (AWS)
Sr. Data Engineer, Strategic Customer Engagements (AWS)

Amazon Inc. • Seattle (WA)

On-site
USD 155,000 - 209,000
Data Engineer, Marketing Tech BI, Stores Finance Analytics & Insights
Data Engineer, Marketing Tech BI, Stores Finance Analytics & Insights

Amazon • Seattle (WA)

On-site
USD 132,000 - 179,000
Health insurance
401(k) matching
Paid time off