Mid Data Engineer - Python, Spark & GCP Data Pipelines

Capco

Kraków

Hybrid

PLN 180,000 - 240,000

Full time

38 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Capco Poland is seeking a Mid Data Engineer to join our growing data engineering team and contribute to scalable data solutions for financial clients.

You will work with modern data technologies and cloud platforms, developing and maintaining data pipelines, processing large datasets, and supporting enterprise-scale data solutions.

This role is ideal for someone with hands-on experience in Python, Spark, Hadoop, Linux, and Google Cloud Platform while collaborating on international projects.

Qualifications

  • 2–4+ years of commercial experience in Data Engineering or a similar role.
  • Hands-on Python programming experience.
  • Experience with Apache Spark and building data processing jobs.
  • Experience with Hadoop or distributed data processing ecosystems.
  • Strong Linux knowledge and CLI proficiency.
  • Commercial experience with Google Cloud Platform (GCP).
  • Understanding of ETL/ELT, data pipelines, and transformation concepts.
  • Proficiency in SQL and relational data concepts.
  • Familiarity with Git and modern software development practices.
  • Ability to work in an Agile environment with distributed teams.
  • Good communication in English (minimum B2).

Responsibilities

  • Design, develop, and maintain scalable data pipelines and processing solutions.
  • Build data transformation workflows using Python and Spark.
  • Work with large-scale datasets in distributed environments like Hadoop.
  • Develop cloud-based data solutions on Google Cloud Platform.
  • Create reliable ingestion processes from multiple source systems.
  • Implement data transformations, validation, and quality checks.
  • Troubleshoot pipelines and optimize performance.
  • Work with Linux-based environments, scripting, and deployment tasks.
  • Collaborate with data engineers, architects, and analysts to translate requirements.
  • Participate in code reviews and follow best practices in data engineering.
  • Document data flows, dependencies, configurations, and operations.
  • Support deployment, testing, stabilization, and maintenance of data solutions.

Skills

Python
Apache Spark
Hadoop
Linux
GCP
SQL
English (B2)
Git
Agile
Data pipelines

Tools

Google Cloud Platform (GCP)
Apache Airflow
Docker/Kubernetes
BigQuery
Dataproc/Dataflow

Job description

Capco Poland is seeking a Mid Data Engineer to join our growing data engineering team and contribute to scalable data solutions for financial clients.

You will work with modern data technologies and cloud platforms, developing and maintaining data pipelines, processing large datasets, and supporting enterprise-scale data solutions.

This role is ideal for someone with hands-on experience in Python, Spark, Hadoop, Linux, and Google Cloud Platform while collaborating on international projects.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Engineer
Data Engineer

Capco • Kraków

Hybrid
PLN 180,000 - 240,000
Senior Data Engineer – GCP, Kraków, Flexible B2B
Senior Data Engineer – GCP, Kraków, Flexible B2B

Capco • Kraków

On-site
PLN 140,000 - 210,000
Mid/Senior Data Engineer (Kraków/GCP)
Mid/Senior Data Engineer (Kraków/GCP)

Capco • Kraków

On-site
PLN 140,000 - 210,000
Cloud Data Engineer – GCP, Spark & Pipelines
Cloud Data Engineer – GCP, Spark & Pipelines

HireHi • Kraków

On-site
PLN 180,000 - 280,000
B2B контракт
AI и автоматизация проектов
Senior Data Engineering Lead - GCP, Spark & Pipelines
Senior Data Engineering Lead - GCP, Spark & Pipelines

GFT Poland • Kraków

Hybrid
PLN 300,000 - 460,000
Competitive salary
Hybrid work model
Training opportunities
Senior GCP Data Engineer - Data Lakes & Pipelines
Senior GCP Data Engineer - Data Lakes & Pipelines

Commit • Warszawa

On-site
PLN 230,000 - 320,000
Cloud Data Engineer: Build Scalable Pipelines & Data Solutions
Cloud Data Engineer: Build Scalable Pipelines & Data Solutions

Capgemini • Wrocław

On-site
PLN 140,000 - 210,000
Company car
Annual bonus
Private medical care
Cloud Data Engineer — Build Scalable ETL Pipelines
Cloud Data Engineer — Build Scalable ETL Pipelines

Capgemini • Polska

Hybrid
PLN 180,000 - 260,000
Company car
Yearly bonus
Medicover private medical care
+8
Data Engineer: Spark, Python & Cloud Data Pipelines
Data Engineer: Spark, Python & Cloud Data Pipelines

Enfint • Łódź

On-site
PLN 102,000 - 122,000
Private medical care
Kafeteria MyBenefit credits
Multisport membership
+3
GCP Data Engineer: Cloud Pipelines, BigQuery & CI/CD
GCP Data Engineer: Cloud Pipelines, BigQuery & CI/CD

Capgemini • Województwo pomorskie

Hybrid
PLN 180,000 - 240,000
Medical care (Medicover)
Private life insurance
Multisport card
+2