Senior Data Scientist: AI-Driven Data Pipelines & Spark

BELVEDERE

Kraków

Hybrid

PLN 180,000 - 260,000

Part time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

BELVEDERE in Krakow is seeking a Senior Data Scientist to design, build, and operate scalable batch and streaming data pipelines. You will work with distributed processing frameworks and AI-assisted development tools to deliver cloud-based data solutions and governance-enabled data workflows.

The role involves collaborating with Analytics and Platform Engineers, leveraging GCP services, and driving data quality and BI use cases. Hybrid work setup in Krakow is offered for a six-month contract.

Qualifications

  • 4-6 years of experience in data engineering or related field.
  • Strong experience with Apache Airflow and distributed data processing frameworks like Apache Spark.
  • Proficiency in Python and SQL for data pipeline development.
  • Mandatory familiarity with AI-assisted coding tools for day-to-day development.
  • Experience with GCP data services for managing cloud-based data solutions.
  • Knowledge of data transformation tools such as dbt.
  • Experience with streaming systems for real-time data processing.
  • Exposure to data quality or governance tooling.
  • Familiarity with analytics or business intelligence use cases.
  • Understanding of Apache Beam for batch and streaming data processing.

Responsibilities

  • Design and build scalable batch and streaming data pipelines.
  • Implement distributed data processing using Apache Spark and Apache Beam.
  • Orchestrate workflows with Apache Airflow to ensure efficient data management.
  • Develop and maintain data transformations using SQL and dbt.
  • Collaborate with Analytics Engineers and Platform Engineers on project execution.
  • Enhance data processing capabilities through AI-driven workflows.
  • Support business intelligence needs through analytics-driven use cases.
  • Improve data integrity using quality and governance tooling.

Skills

Python
SQL
AI-assisted coding tools
GCP familiarity

Tools

Apache Spark
Apache Beam
Apache Airflow
dbt

Job description

BELVEDERE in Krakow is seeking a Senior Data Scientist to design, build, and operate scalable batch and streaming data pipelines. You will work with distributed processing frameworks and AI-assisted development tools to deliver cloud-based data solutions and governance-enabled data workflows.

The role involves collaborating with Analytics and Platform Engineers, leveraging GCP services, and driving data quality and BI use cases. Hybrid work setup in Krakow is offered for a six-month contract.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Scientist
Senior Data Scientist

BELVEDERE • Kraków

Hybrid
PLN 180,000 - 260,000
Senior Data Scientist: Python ML Infra & GPU Pipelines
Senior Data Scientist: Python ML Infra & GPU Pipelines

BELVEDERE • Kraków

Hybrid
PLN 240,000 - 280,000
Senior Data Engineer - AI Pipelines & Snowflake
Senior Data Engineer - AI Pipelines & Snowflake

IBM Computing • Kraków

On-site
PLN 240,000 - 400,000
Senior Data Engineering Lead - GCP, Spark & Pipelines
Senior Data Engineering Lead - GCP, Spark & Pipelines

GFT Poland • Kraków

Hybrid
PLN 300,000 - 460,000
Competitive salary
Hybrid work model
Training opportunities
Senior Data Engineer: Build Scalable Pipelines & AI Tools
Senior Data Engineer: Build Scalable Pipelines & AI Tools

Sii Poland • Warszawa

On-site
PLN 210,000 - 270,000
Great Place to Work
Profit sharing
Medical care
+2
Senior Data Platform Lead: GCP, PySpark & AI Pipelines
Senior Data Platform Lead: GCP, PySpark & AI Pipelines

JobCubby • Kraków

Hybrid
PLN 194,000 - 298,000
Hybrid work in Kraków
Medical, sport, lunch subsidy
Online training
+1
Senior Databricks Data Engineer — Scalable Pipelines + AI
Senior Databricks Data Engineer — Scalable Pipelines + AI

Capgemini • Warszawa

On-site
PLN 90,000 - 130,000
Medicover medical care
Private life insurance
Sports card
+3
Remote Senior Data Engineer — Data Pipelines & Analytics
Remote Senior Data Engineer — Data Pipelines & Analytics

PwC Polska • Kraków

Remote
PLN 180,000 - 280,000
Hybrid working model
Mentoring from experienced colleagues
Medical care package
+2
Senior Databricks Data Engineer — Drive Data Pipelines
Senior Databricks Data Engineer — Drive Data Pipelines

Capgemini • Kraków

On-site
PLN 120,000 - 190,000
Medicover health care
Private life insurance
Sports card
+1
Hybrid Data Engineer – GCP, Real-Time Pipelines
Hybrid Data Engineer – GCP, Real-Time Pipelines

IG KnowHow • Kraków

Hybrid
Tailored development programmes
Mentoring opportunities
Clear career progression
+1