Senior Data Pipeline Engineer — Python, Spark & Kafka

Jobtailor

Kentucky

On-site

USD 110,000 - 140,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Jobtailor in Kentucky is seeking a Data Engineer to design, build, and maintain scalable data pipelines and ETL/ELT processes, leveraging Python, PySpark, and Java to transform large datasets.

You will ingest structured and unstructured data via REST APIs, streaming platforms, and batch jobs, work with Kafka and Hadoop, and ensure data quality and governance in an Agile team environment.

Qualifications

  • Bachelor's degree required and 4+ years in Data Engineering or Big Data development.
  • Hands-on experience with Python, PySpark, Java, Kafka, Hadoop, REST APIs, and SQL.
  • Strong analytical and problem-solving abilities.
  • Ability to work effectively in both team-oriented and independent environments.

Responsibilities

  • Design, develop, and maintain scalable data pipelines and ETL/ELT processes.
  • Build and optimize large-scale data processing solutions using Python, PySpark, and Java.
  • Develop data ingestion frameworks for structured and unstructured data sources.
  • Integrate data from various systems through REST APIs, streaming platforms, and batch processing.
  • Work with large datasets in distributed computing environments.
  • Develop and support real-time and batch data processing solutions using Kafka and Hadoop ecosystem technologies.
  • Implement scalable streaming data pipelines and event-driven architectures.
  • Monitor and optimize data workflows for performance and reliability.
  • Write complex SQL queries for data extraction, transformation, and analysis.
  • Design and optimize database schemas and data models.
  • Ensure data quality, consistency, and governance standards are maintained.
  • Participate in Agile ceremonies including sprint planning, daily stand-ups, backlog grooming, and retrospectives.
  • Use Jira for project tracking, issue management, and sprint execution.
  • Collaborate with Data Architects, Data Scientists, Business Analysts, and Application Development teams.
  • Work independently as well as within a collaborative team environment.
  • Troubleshoot and resolve data-related issues across environments.
  • Perform root cause analysis and implement long-term solutions.
  • Support production deployments and ongoing maintenance activities.

Skills

Python
PySpark
Java
SQL
ETL/ELT
Data modeling
Distributed computing

Education

Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field

Tools

Kafka
Hadoop
REST APIs
Jira

Job description

Jobtailor in Kentucky is seeking a Data Engineer to design, build, and maintain scalable data pipelines and ETL/ELT processes, leveraging Python, PySpark, and Java to transform large datasets.

You will ingest structured and unstructured data via REST APIs, streaming platforms, and batch jobs, work with Kafka and Hadoop, and ensure data quality and governance in an Agile team environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer: Real-Time Data Pipelines
Senior Data Engineer: Real-Time Data Pipelines

Jobtailor • California (MO)

On-site
USD 140,000 - 190,000
Lead Data Engineer: Python & AWS for Real-Time Data
Lead Data Engineer: Python & AWS for Real-Time Data

Jobtailor • Massachusetts

On-site
USD 140,000 - 200,000
Senior Data & AI Engineer: Cloud Pipelines & ML
Senior Data & AI Engineer: Cloud Pipelines & ML

Jobtailor • Plano (TX)

On-site
USD 120,000 - 160,000
Senior Data Engineer - Spark/PySpark & Data Platform Migrations
Senior Data Engineer - Spark/PySpark & Data Platform Migrations

Jobtailor • Santa Clara (CA)

On-site
USD 150,000 - 230,000
Principal Data Engineer - Data Platform & AI Pipelines
Principal Data Engineer - Data Platform & AI Pipelines

Jobtailor • New York (NY)

On-site
USD 150,000 - 190,000
Senior Data Engineer: Analytics, Pipelines & BI
Senior Data Engineer: Analytics, Pipelines & BI

Jobtailor • Arizona

On-site
USD 90,000 - 115,000
Senior Data Platform Engineer – Spark, Python & Hadoop
Senior Data Platform Engineer – Spark, Python & Hadoop

Jobtailor • Alabama

On-site
USD 85,000 - 125,000
Senior Data Engineer, Real-Time Pipelines & CDP Architect
Senior Data Engineer, Real-Time Pipelines & CDP Architect

Jobtailor • Seattle (WA)

On-site
USD 120,000 - 190,000
Senior Data Engineer — Big Data Platform & Scalable Pipelines
Senior Data Engineer — Big Data Platform & Scalable Pipelines

Jobtailor • Alabama

On-site
USD 120,000 - 150,000
Senior Data Engineer: Scalable Pipelines & Gen AI
Senior Data Engineer: Scalable Pipelines & Gen AI

Jobtailor • Chicago (IL)

On-site
USD 130,000 - 180,000