Senior Data Platform Engineer - Cloud & Pipelines

Pioneering Intelligence

Cambridge (MA)

On-site

USD 128,000 - 176,000

Full time

4 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Healthcare coverage
Annual incentive program
Retirement benefits

Job summary

Flagship Pioneering is seeking a Lead Data Engineer to spearhead modernization of data infrastructure, pipelines, and platforms. You will work with I&O, Lab IT, and scientific teams to turn evolving needs into secure, scalable data solutions.

Strong AWS, Python, and SQL skills are essential, as is experience with Dagster, dbt, and Spark. You will mentor engineers, define standards, and leverage generative AI to accelerate delivery, all while ensuring reliability and performance of data

Qualifications

  • 6+ years in data engineering, preferably in life sciences or healthcare.
  • Proven cloud-native data architectures on AWS.
  • Strong Python and SQL skills for pipelines and automation.
  • Experience with modern data frameworks and tooling (Dagster, dbt, Spark).
  • Experience with Git, CI/CD, containerized workloads (Docker, ECS).

Responsibilities

  • Lead data infrastructure projects across interdependent systems and cloud platforms.
  • Implement data platforms and pipelines; optimize storage solutions.
  • Establish engineering standards for testing, release management, and monitoring.
  • Mentor engineers and drive adoption of generative AI to speed pipeline work.
  • Collaborate with I&O, Lab IT, PI Tech, and data science for architectural alignment.

Skills

Data engineering
Python
SQL
AWS
Generative AI
6+ years experience
Cloud data architectures
Data pipelines
Leadership
Cross-functional collaboration
Troubleshooting

Tools

Dagster
dbt
Spark
Iceberg
Lake Formation
Athena
Glue
ECS
Docker
Git
CI/CD
CDK
Spark SQL

Job description

Flagship Pioneering is seeking a Lead Data Engineer to spearhead modernization of data infrastructure, pipelines, and platforms. You will work with I&O, Lab IT, and scientific teams to turn evolving needs into secure, scalable data solutions.

Strong AWS, Python, and SQL skills are essential, as is experience with Dagster, dbt, and Spark. You will mentor engineers, define standards, and leverage generative AI to accelerate delivery, all while ensuring reliability and performance of data

Get your free, confidential resume review.
or drag and drop your file here.