Onsite Data Engineer — Spark, Kafka & AWS Pipelines

Pyramid Consulting, Inc

North Carolina

On-site

USD 69,000 - 83,000

Full time

8 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
401(k) plan
Paid sick leave

Job summary

Pyramid Consulting, Inc. is seeking a Data Engineer for a 12+ month contract opportunity located onsite in the United States. The role involves designing and maintaining scalable data pipelines using Python, Spark, Kafka, Airflow, and AWS services, with emphasis on Databricks and data warehousing.

Responsibilities include building ETL/ELT pipelines, real-time and batch processing, and collaboration with data scientists and stakeholders to ensure data quality and performance across platforms.

Qualifications

  • Must have experience designing scalable data pipelines with Python and Spark.
  • Experience with Kafka, Airflow, and AWS data services.
  • Experience with Databricks and data warehousing.
  • Onsite from Day 1 in United States.

Responsibilities

  • Design, develop, and maintain scalable data pipelines using Python, PySpark, and Apache Spark.
  • Build and optimize ETL/ELT pipelines to ingest, transform, and process large volumes of structured and unstructured data.
  • Develop real-time and batch data pipelines using Apache Kafka and Airflow for data ingestion and workflow orchestration.
  • Build and manage cloud-based data solutions using AWS S3, Glue, EMR, and Redshift.
  • Develop and optimize data processing workloads using Databricks and PySpark.
  • Write complex SQL queries and optimize data processing for performance and scalability.
  • Design and maintain data warehouses, data models, and scalable data architectures.
  • Develop reusable data ingestion and transformation frameworks to support analytics and reporting requirements.
  • Implement workflow scheduling, dependency management, monitoring, and failure handling using Airflow and other orchestration tools.
  • Ensure data quality, accuracy, reliability, and consistency across data pipelines and downstream systems.
  • Collaborate with data scientists, analysts, architects, and business stakeholders to understand data requirements and deliver scalable data solutions.
  • Troubleshoot pipeline failures and performance issues and continuously improve the reliability and efficiency of data platforms.

Skills

Python
Apache Spark
SQL
Kafka
Airflow
AWS
Databricks
ETL/ELT pipelines
data warehousing
data modeling
orchestration tools

Tools

PySpark

Job description

Pyramid Consulting, Inc. is seeking a Data Engineer for a 12+ month contract opportunity located onsite in the United States. The role involves designing and maintaining scalable data pipelines using Python, Spark, Kafka, Airflow, and AWS services, with emphasis on Databricks and data warehousing.

Responsibilities include building ETL/ELT pipelines, real-time and batch processing, and collaboration with data scientists and stakeholders to ensure data quality and performance across platforms.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Onsite Hadoop & Spark Data Engineer: ETL, Kafka, Airflow
Onsite Hadoop & Spark Data Engineer: ETL, Kafka, Airflow

Pyramid Consulting, Inc • North Carolina

On-site
USD 69,000 - 83,000
Health insurance
401(k) plan
Paid sick leave
Data Engineer
Data Engineer

Pyramid Consulting, Inc • North Carolina

On-site
USD 69,000 - 83,000
Health insurance
401(k) plan
Paid sick leave
Senior Data Engineer: Pipelines, Cloud & Data Lakes
Senior Data Engineer: Pipelines, Cloud & Data Lakes

hackajob • Plano (TX)

On-site
USD 110,000 - 150,000
Health care coverage
On-site wellness centers
Retirement plan
+4
Senior Data Engineer — Onsite Cloud & Databricks ETL Lead
Senior Data Engineer — Onsite Cloud & Databricks ETL Lead

JPS Tech Solutions • Seattle (WA)

On-site
USD 170,000 - 210,000
Databricks Data Engineer - Spark, ETL & Cloud Pipelines
Databricks Data Engineer - Spark, ETL & Cloud Pipelines

Smart IT Frame LLC • Reston (VA)

On-site
USD 90,000 - 120,000
Senior Data Engineer — AWS Cloud Pipelines (Remote)
Senior Data Engineer — AWS Cloud Pipelines (Remote)

Attain Talent • United States

Hybrid
USD 110,000 - 140,000
Remote Work (Hybrid)
Medical, Dental, Vision
401(k) with matching
+2
Data Engineer
Data Engineer

JPS Tech Solutions • Seattle (WA)

On-site
USD 170,000 - 210,000
AWS Python Developer with Pyspark Newark, NJ, New Jersey
AWS Python Developer with Pyspark Newark, NJ, New Jersey

Polarits • Newark (NJ)

On-site
USD 140,000 - 190,000
Senior PySpark Data Engineer
Senior PySpark Data Engineer

Covetus • Irving (TX)

On-site
USD 100,000 - 130,000
AWS Data Engineer
AWS Data Engineer

Interon IT Solutions • Chantilly (VA)

On-site
USD 125,000 - 170,000