Lead Data Engineer – AWS, Databricks & ETL/ELT

ScolerTec Inc

United States

On-site

USD 140,000 - 190,000

Full time

2 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

ScolerTec Inc. seeks a Lead Data Engineer to drive a large-scale cloud data modernization and governance program in the US.

You will lead design, development, testing, and maintenance of secure, scalable data engineering components and platform services for operational data, reporting, analytics, and AI/ML use cases. You will collaborate with architecture, governance, migration, and DevSecOps teams to deliver auditable cloud-based data solutions using AWS and Databricks.

Qualifications

  • 7+ years of specialized data engineering experience, including leadership of enterprise-scale engineering efforts.
  • Strong hands-on experience with AWS data and compute services.
  • Strong hands-on production experience with Databricks, Spark/PySpark, and Delta Lake/lakehouse architectures.
  • Strong proficiency in SQL, Python, ETL/ELT, and data transformation.
  • Hands-on experience building enterprise-scale batch and near-real-time data pipelines.
  • Experience with AWS data services such as S3, Glue, Redshift, Kinesis, Lambda, Athena, or related services.
  • Experience with orchestration tools and streaming/near-real-time ingestion.
  • Experience with CI/CD, secure software engineering, and Infrastructure as Code.
  • Experience with data governance, metadata, lineage, data quality, and classification standards.
  • Experience leading and mentoring Data Engineers.
  • Federal or regulated-environment experience preferred.

Responsibilities

  • Lead development of scalable cloud-based data platforms supporting data lake, lakehouse, and data warehouse architectures.
  • Design and develop enterprise ETL/ELT and ingestion pipelines for batch and near-real-time workloads.
  • Build and optimize data engineering solutions using Databricks, Spark/PySpark, Delta Lake, Python, and SQL.
  • Implement secure and auditable integrations across legacy, cloud, and external systems.
  • Lead implementation of operational databases, document storage, schemas, APIs, and data services.
  • Provide technical direction and oversight to Data Engineers, including design and code reviews.
  • Implement Infrastructure as Code for database and platform provisioning and configuration.
  • Implement data quality checks, validation, error handling, metadata, lineage, and governance-aligned structures.
  • Optimize pipeline performance, Spark workloads, query performance, scalability, and cloud cost.
  • Integrate pipelines with Government-provided CI/CD processes.
  • Support incident resolution, RCA, migration, archival, failover testing, and DR activities.
  • Maintain technical documentation, architecture artifacts, runbooks, and audit-ready deliverables.

Skills

Data engineering leadership
SQL
Python
ETL/ELT
Spark/PySpark
Delta Lake
Cloud data platforms
CI/CD
Data governance
Mentoring

Tools

Databricks
Spark
Delta Lake
Unity Catalog
Workflows
Auto Loader
Terraform
Kinesis

Job description

ScolerTec is seeking multiple Lead Data Engineers to support a large-scale cloud data modernization and governance program. This role will lead the design, development, testing, and maintenance of secure, scalable data engineering components and platform services supporting operational data, reporting, analytics, and AI/ML use cases. The Lead Data Engineer will work closely with architecture, governance, migration, and DevSecOps teams to deliver high-quality, auditable, and performant cloud-based data solutions using AWS and Databricks.

Key Responsibilities
  • Lead development of scalable cloud-based data platforms supporting data lake, lakehouse, and data warehouse architectures.
  • Design and develop enterprise ETL/ELT and ingestion pipelines for batch and near-real-time workloads.
  • Build and optimize data engineering solutions using Databricks, Spark/PySpark, Delta Lake, Python, and SQL.
  • Implement secure and auditable integrations across legacy, cloud, and external systems.
  • Lead implementation of operational databases, document storage, schemas, APIs, and data services.
  • Provide technical direction and oversight to Data Engineers, including design and code reviews.
  • Implement Infrastructure as Code for database and platform provisioning and configuration.
  • Implement data quality checks, validation, error handling, metadata, lineage, and governance-aligned structures.
  • Optimize pipeline performance, Spark workloads, query performance, scalability, and cloud cost.
  • Integrate pipelines with Government-provided CI/CD processes.
  • Support incident resolution, RCA, migration, archival, failover testing, and DR activities.
  • Maintain technical documentation, architecture artifacts, runbooks, and audit-ready deliverables.
Qualifications
  • 7+ years of specialized data engineering experience, including leadership of enterprise-scale engineering efforts.
  • Strong hands-on experience with AWS data and compute services.
  • Strong hands-on production experience with Databricks, Spark/PySpark, and Delta Lake/lakehouse architectures.
  • Strong proficiency in SQL, Python, ETL/ELT, and data transformation.
  • Hands-on experience building enterprise-scale batch and near-real-time data pipelines.
  • Experience with AWS data services such as S3, Glue, Redshift, Kinesis, Lambda, Athena, or related services.
  • Experience with orchestration tools and streaming/near-real-time ingestion.
  • Experience with CI/CD, secure software engineering, and Infrastructure as Code.
  • Experience with data governance, metadata, lineage, data quality, and classification standards.
  • Experience leading and mentoring Data Engineers.
  • Federal or regulated-environment experience preferred.
Preferred
  • Databricks experience with Unity Catalog, Workflows, Auto Loader, SQL Warehouses, Delta Live Tables/Lakeflow, or governed Gold-layer datasets.
  • Informatica or another enterprise integration/ETL platform.
  • Terraform or equivalent Infrastructure as Code.
  • Kafka or Kinesis streaming experience.
Preferred Certifications
  • AWS Certified Solutions Architect – Professional
  • AWS Certified Data Analytics – Specialty
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Lead Data Engineer - AWS & Databricks Platform
Lead Data Engineer - AWS & Databricks Platform

ScolerTec Inc • United States

On-site
USD 140,000 - 190,000
Lead Data Engineer
Lead Data Engineer

Strategic Staffing Solutions • Charlotte (NC)

On-site
USD 124,000 - 138,000
Lead Data Engineer
Lead Data Engineer

Recru • Houston (TX)

On-site
USD 130,000 - 150,000
Data Lead
Data Lead

Prodapt • Richardson (TX)

On-site
USD 120,000 - 160,000
Lead Data Engineer with Databricks
Lead Data Engineer with Databricks

Univedge Consulting LLC • St. Louis (MO)

On-site
USD 120,000 - 180,000
Data Lead
Data Lead

Prodapt Solutions Private Limited • Richardson (TX)

On-site
USD 140,000 - 190,000
Senior Data Engineer - Databricks
Senior Data Engineer - Databricks

DATAECONOMY Inc • New Jersey

On-site
USD 130,000 - 180,000
Senior Data Engineer
Senior Data Engineer

Compunnel, Inc. • Charlotte (NC)

On-site
USD 120,000 - 150,000
Lead Data Engineer
Lead Data Engineer

Recru • Spring (TX)

On-site
USD 130,000 - 170,000
Sr. Databricks Engineer - Hybrid NYC
Sr. Databricks Engineer - Hybrid NYC

SOMERSET STAFFING • New York (NY)

On-site
USD 140,000 - 190,000