Data Engineer II, Data Engineer,Data Center Capacity Delivery

Amazon

Seattle, Northern (WA, KY)

Hybrid

USD 132,000 - 179,000

Full time

11 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Health insurance
401(k) matching
Paid time off
Parental leave

Job summary

Amazon is seeking a Data Engineer to support data lake and warehouse systems within AWS Data Services. You’ll join a diverse team of engineers, construction specialists, and security experts to deliver scalable data solutions using a serverless architecture, primarily with AWS components including Redshift, S3, Glue, EMR, and Lambda.

You will design and operate ETL pipelines, ensure data quality, and implement governance while collaborating with Data Scientists, PMs, and SDEs to meet data

Qualifications

  • Bachelor's degree in a technical field and 3+ years of data engineering experience.
  • Experience with data modeling, warehousing and ETL pipelines.
  • Proficient in SQL and programming languages such as Python or Java.
  • Strong knowledge of distributed data systems, batch and streaming architectures.
  • Experience with AWS data services and modern data tooling.

Responsibilities

  • Design scalable, fault-tolerant data pipelines using AWS tech and internal tools.
  • Collaborate with cross-functional teams to define data requirements.
  • Automate deployment with CI/CD pipelines and streamline maintenance.
  • Ensure data quality through validation, cleansing, and deduplication.
  • Implement data governance, access control, encryption, retention, and audits.
  • Continuously improve pipelines and stay current with new tech.
  • Build scalable data platforms supporting analytics and self-service data products.
  • Write high-quality code interfacing with critical services and APIs.
  • Create end-to-end pipelines consolidating data from multiple sources.

Skills

Python
SQL
Data modeling
ETL pipelines
AWS data services
Spark
Distributed systems
CI/CD
Data quality
Communication

Education

Bachelor's degree

Tools

Apache Spark
EMR
Kinesis
Glue
Airflow

Job description

Job ID: 10542353 | Amazon Data Services, Inc.

AWS Data Center Capacity Delivery (DCCD) is looking for a Data Engineer to support data center construction globally. We work on the most challenging problems, with thousands of variables impacting the data center delivery — and we’re looking for talented people who want to help.

You’ll join a diverse team of software, hardware, and network engineers, construction specialists, security experts, operations managers, and other vital roles. You’ll collaborate with people across AWS to help us deliver the highest standards for safety and security while providing seemingly infinite capacity at the lowest possible cost for our customers. You’ll experience an inclusive culture that welcomes bold ideas and empowers you to own them to completion.

We’re looking for Data Engineer to help us grow our Data Lake and Data Warehouse Systems, which is being built using a serverless architecture, with 100% native AWS components including Redshift Spectrum, Athena, S3, Lambda, Glue, EMR, Kinesis, SNS, CloudWatch and more! We own a world-class data lake that is used to drive multi-billion dollar decisions on a regular cadence and we're looking to improve on filling the lake quickly, with as little human intervention needed and democratize the data in the lake.

Our Data Engineers build the ETL and analytics solutions for our internal customers to answer questions with data and drive critical improvements for the business. Our Data Engineers use best practices in software engineering, data management, data storage, data compute, and distributed systems. We are passionate about solving business problems with data!

Key job responsibilities
  • Design and implement scalable, fault-tolerant data pipelines using AWS technologies and internal Amazon tools to extract, transform, and load data from multiple sources leveraging and implementing AI solutions as required.
  • Collaborate cross-functionally with BIEs, Data Scientists, PMs, and SDEs to understand data requirements and deliver customized data solutions.
  • Automate infrastructure deployment with CI/CD pipelines and ensure streamlined processes for deployment and maintenance.
  • Ensure data quality through robust validation, cleansing, and deduplication techniques.
  • Implement data governance standards, including access control, encryption, data retention, deletion policies, and audit mechanisms to ensure compliance and security.
  • Continuously improve and optimize data pipelines and infrastructure, staying up to date with emerging technologies and implementing automation and monitoring tools.
  • Build a scalable and reliable data platform supporting analytics for intuitive, self-service data products.
  • Write high quality code and build scalable applications that interface with critical services and APIs to extract and process unstructured data
  • Work with a range of data technologies, including Python, EMR, Spark, Iceberg, Airflow, and many AWS data services like Glue, Athena, Redshift to create end-to-end pipelines that consolidate data from disparate systems.
About the team

DCCD -CAT is a central data and analytics team within the DCCD tooling org that plays a pivotal role in supporting analytics for data center construction space suporting cost,controls and commissioning domains. We own data platform, reporting, dashboards, measurement, and analytical solutions for DCCD org.

Basic Qualifications
  • - Bachelor's degree
  • - 3+ years of data engineering experience
  • - Experience with data modeling, warehousing and building ETL pipelines
  • - Experience with SQL
  • - Knowledge of professional software engineering & best practices for full software development life cycle, including coding standards, software architectures, code reviews, source control management, continuous deployments, testing, and operational excellence
  • - Knowledge of distributed systems as it pertains to data storage and computing
  • - Knowledge of batch and streaming data architectures like Kafka, Kinesis, Flink, Storm, Beam
  • - Experience as a data engineer or related specialty (e.g., software engineer, business intelligence engineer, data scientist) with a track record of manipulating, processing, and extracting value from large datasets
  • - Experience in at least one modern scripting or programming language, such as Python, Java, Scala, or NodeJS
  • - Experience with Apache Spark / Elastic Map Reduce
Preferred Qualifications
  • - Experience with AWS technologies like Redshift, S3, AWS Glue, EMR, Kinesis, FireHose, Lambda, and IAM roles and permissions
  • - Experience with non-relational databases / data stores (object storage, document or key-value stores, graph databases, column-family databases)
  • - Master's degree in computer science, engineering, analytics, mathematics, statistics, IT or equivalent
  • - Experience programming with at least one modern language such as C++, C#, Java, Python, Golang, PowerShell, Ruby
  • - Experience building/operating highly available, distributed systems of data extraction, ingestion, and processing of large data sets
  • - Experience in translating business needs into detailed feature requirements

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

Preferred Qualifications
  • - Experience with AWS technologies like Redshift, S3, AWS Glue, EMR, Kinesis, FireHose, Lambda, and IAM roles and permissions
  • - Experience with non-relational databases / data stores (object storage, document or key-value stores, graph databases, column-family databases)
  • - Master's degree in computer science, engineering, analytics, mathematics, statistics, IT or equivalent
  • - Experience programming with at least one modern language such as C++, C#, Java, Python, Golang, PowerShell, Ruby
  • - Experience building/operating highly available, distributed systems of data extraction, ingestion, and processing of large data sets
  • - Experience in translating business needs into detailed feature requirements

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

The base salary range for this position is listed below. Your Amazon package will include sign‑on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits .

  • health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage)
  • 401(k) matching
  • paid time off
  • parental leave

The base salary range for this position is listed below. Your Amazon package will include sign‑on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits .

USA, WA, Seattle - 132,100.00 - 178,800.00 USD annually

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer II, Data Engineer,Data Center Capacity Delivery
Data Engineer II, Data Engineer,Data Center Capacity Delivery

Socket.dev • Seattle (WA)

On-site
USD 132,000 - 179,000
Health insurance
401(k) matching
Paid time off
+1
Data Engineer, AWS DC Central Operations
Data Engineer, AWS DC Central Operations

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 101,000 - 160,000
Health insurance
401(k) matching
Paid time off
+2
Data Engineer, Amazon Customer Service
Data Engineer, Amazon Customer Service

Socket.dev • Seattle (WA)

On-site
USD 132,000 - 179,000
Data Engineer, WW Ops Finance - S&A
Data Engineer, WW Ops Finance - S&A

Amazon • Factoria (WA)

On-site
USD 132,000 - 179,000
Health insurance
401(k) matching
Paid time off
Data Engineer, Deal Tooling and Insights, Strategic Customer Engagements
Data Engineer, Deal Tooling and Insights, Strategic Customer Engagements

Amazon Web Services (AWS) • Arlington (VA)

On-site
USD 132,000 - 179,000
Environmental Data Engineer, AWS Environmental
Environmental Data Engineer, AWS Environmental

Amazon Web Services (AWS) • Herndon (VA)

On-site
USD 132,000 - 179,000
Health insurance
401(k) matching
Paid time off
+1
Environmental Data Engineer, AWS Environmental
Environmental Data Engineer, AWS Environmental

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 132,000 - 179,000
Health insurance
401(k) matching
Paid time off
Data Engineer, SPTC
Data Engineer, SPTC

Amazon • Seattle (WA), Northern (KY)

Hybrid
USD 132,000 - 179,000
Health insurance
401(k) matching
Paid time off
+2
Data Engineer II, DBS BI
Data Engineer II, DBS BI

Amazon • Bellevue (WA)

On-site
USD 132,000 - 179,000
Health insurance
401(k) matching
Paid time off
+1
Senior Data Engineer, AWS Analytics Engineering
Senior Data Engineer, AWS Analytics Engineering

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 155,000 - 209,000