Data Engineer (AWS, Spark)

Peregrine Advisors LLC

Washington (District of Columbia)

Hybrid

USD 103,000 - 140,000

Full time

6 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Medical insurance
Dental insurance
Vision insurance
401(k) match
Life insurance
Disability insurance
Unlimited PTO
Certifications support

Job summary

Peregrine Advisors LLC is seeking a Data Engineer in the Washington, DC area to build scalable ETL pipelines on AWS using Spark, Glue, and EMR. You will design S3 data lakes, feed Apache Iceberg tables, and coordinate data across Aurora PostgreSQL and DynamoDB.

The role emphasizes data quality, governance, and scalable automation. You will work in a hybrid setting, interact with engineers and strategists, and contribute to a high-impact federal-ready data platform.

Qualifications

  • 4+ years of data engineering experience.
  • Bachelor's degree.
  • Spark ETL on AWS (Glue, Amazon EMR) in Python and PySpark.
  • S3 data-lake design (Parquet, partitioning, lifecycle) feeding Apache Iceberg tables, Amazon Aurora PostgreSQL, and DynamoDB.
  • Event orchestration (Lambda, Step Functions, SQS/SNS) with secrets management and monitoring.
  • Data quality, validation, and lineage.
  • Infrastructure-as-code (CloudFormation or Terraform).
  • Basic proficiency in writing, PowerPoint, and Excel.

Responsibilities

  • Build ETL pipelines with Spark on AWS.
  • Write processing in Python and PySpark.
  • Design the S3 data-lake and Parquet layout feeding Iceberg tables.
  • Connect PostgreSQL on Aurora and DynamoDB, with Trino for federated SQL.
  • Deliver pipelines that keep data clean and trustworthy at scale.
  • Automate data ingestion and storage design for downstream use.

Skills

Data engineering
Spark ETL
Python
PySpark
S3 data lake
AWS Lambda/Step Functions
Data quality
Terraform

Education

Bachelor's degree

Tools

Glue
Amazon EMR
Aurora PostgreSQL
DynamoDB
Trino
Iceberg

Job description

Description

At Peregrine Advisors, you will build the pipelines that move a federal agency's data from source to platform. But we hire people, not seats. As the work evolves, you will learn new systems and tools, take on greater responsibility, and help develop the firm's capabilities, tools, and lines of business. We move our best to where the hardest problems are. This is a hybrid work arrangement based in the Washington, DC metropolitan area, and the commuting cadence varies by assignment. The initial engagement requires United States citizenship and the ability to obtain a Public Trust determination. We are a data and technology innovation hub and a Benefit Corporation working at the center of the federal government's mission to deliver for client stakeholders and the US public.

Your first project

Your first project will likely have you building ingest, processing, and storage at scale. Depending on the assignment, you may:

  • Build Spark-based extract, transform, and load (ETL) pipelines with Glue, Amazon EMR, Lambda, and Step Functions.
  • Write the processing in Python and PySpark.
  • Design the S3 layer, including Parquet, partitioning, and lifecycle, feeding Apache Iceberg tables.
  • Connect PostgreSQL on Amazon Aurora and DynamoDB, with Trino for federated Structured Query Language (SQL) across them.

You deliver pipelines that keep data clean and trustworthy at the volumes a federal platform runs at. You automate ingestion that used to be handled case by case, and you build the storage design everything downstream depends on.

This is where you start, not the shape of your career here.

Requirements
  • 4+ years of data engineering experience
  • Bachelor's degree
  • Spark ETL on AWS (Glue, Amazon EMR) in Python and PySpark
  • S3 data-lake design (Parquet, partitioning, lifecycle) feeding Apache Iceberg tables, Amazon Aurora PostgreSQL, and DynamoDB
  • Event orchestration (Lambda, Step Functions, SQS/SNS) with secrets management and monitoring
  • Data quality, validation, and lineage
  • Infrastructure-as-code (CloudFormation or Terraform)
  • Basic proficiency in writing, PowerPoint, and Excel
Preferred
  • Master's degree in a relevant field
  • Trino or comparable federated SQL across the lake and relational stores
  • Apache Ranger-governed access
  • Legacy ETL migration (for example DataStage)
  • Federal information technology or high-volume data experience
  • Familiarity with AI-assisted developer tooling
Who You Are

You are a data engineer who wants to get better at it, and you know which parts you have not mastered yet. You care as much about whether the data is trustworthy as whether it arrives, and you do your best work alongside people who push you. You experiment, fail, learn, and repeat quickly. You would rather own an outcome than be handed a task.

What You Bring
  • Hands-on data engineering on AWS: Spark ETL (Glue, EMR), Python and PySpark, and S3 data-lake design feeding the platform stores.
  • The reliability craft around it: event orchestration, data quality and lineage, monitoring, and infrastructure-as-code.
  • The judgment to build in a regulated environment where accuracy and auditability are not optional.

You adapt your development workflow as AI tools evolve, using them to help implement, test, and improve the components you own. You give the tools clear context, review and test their output, and remain accountable for the code you deliver.

When we talk

Be prepared to discuss a difficult problem you worked through, the decisions you made, what happened, and what you learned.

Benefits

This is a full-time W-2 position with a salary of $103,000 to $140,000 per year.

Benefits include medical, dental, and vision insurance with the employee premium fully paid and half of dependent premiums; employer-paid life, accidental death, and short-term and long-term disability insurance; a 401(k) matched 100% up to 4% of salary, vesting immediately; unlimited paid time off; and sponsored professional certifications and continuing education.

What We Offer

You will work alongside developers, engineers, data scientists, architects, and strategists, on work ranging from strategy to implementation. We support your development across assignments and clients through extensive onboarding, sponsored professional certifications such as the Data Management Capability Assessment Model (DCAM) and AWS technical certifications, and rotation across functions to expand your skills and perspective. The mission is real, the problems are hard, and you own what you ship.

What We Commit To

As a Benefit Corporation, our commitment runs three ways: real, measurable value for our clients; government that works better for the public; and a team that makes everyone in it better.

We hire people who want to help build the firm, not just work at it. If that is you, apply.

Peregrine Advisors is an equal opportunity employer.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Engineer (AWS, Spark)
Data Engineer (AWS, Spark)

Peregrine Advisors • Washington

On-site
USD 120,000 - 160,000
Health insurance
401(k) match
Unlimited PTO
+1
Data Architect (AWS, Data Lake)
Data Architect (AWS, Data Lake)

Peregrine Advisors LLC • Washington

Hybrid
USD 120,000 - 180,000
Medical coverage
Dental coverage
Vision coverage
+3
Database Developer (SQL)
Database Developer (SQL)

Peregrine Advisors LLC • Washington

Hybrid
USD 120,000 - 140,000
Medical, dental, vision insurance
Life & disability insurance
401(k) match
+2
Data Analyst (SQL, AWS)
Data Analyst (SQL, AWS)

Peregrine Advisors • Washington

On-site
USD 75,000 - 110,000
401(k) matched
Paid time off
Professional certifications
+1
Junior Database Developer
Junior Database Developer

Peregrine Advisors LLC • Washington

Hybrid
USD 80,000 - 95,000
Medical, Dental, Vision
401(k) match
Unlimited PTO
+2
Data Analyst (SQL, AWS)
Data Analyst (SQL, AWS)

Peregrine Advisors LLC • Washington

Hybrid
USD 90,000 - 130,000
Medical, dental, and vision fully paid
401(k) matched 100% up to 4%
Unlimited paid time off
+1
Junior Database Developer
Junior Database Developer

Peregrine Advisors • Washington

Hybrid
USD 80,000 - 95,000
DevSecOps / Platform Engineer
DevSecOps / Platform Engineer

Peregrine Advisors LLC • Washington

Hybrid
USD 120,000 - 150,000
Health insurance
401(k) match
Paid time off
Junior DevSecOps Engineer (CI/CD)
Junior DevSecOps Engineer (CI/CD)

Peregrine Advisors LLC • Washington

Hybrid
USD 80,000 - 95,000
Medical, dental, vision insurance
Employer-paid life, AD&D and STD/ LTD
401(k) with 100% match up to 4%
+2
Data Scientist, Python and AWS
Data Scientist, Python and AWS

Remote Jobs • United States

Hybrid
USD 78,000 - 147,000
Medical/Dental/Vision Insurance
Employer-paid life & disability
401(k) matching
+2