AI Data Engineer (ML Data Pipelines)

Zohorecruit

Polska

Hybrid

PLN 190,000 - 260,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Remote work

Job summary

Zohorecruit is seeking an AI Data Engineer to design and build production-grade data pipelines powering machine learning systems. You will craft scalable ingestion, transformation, and feature engineering workflows for model training, evaluation and real-time inference.

You will collaborate with data scientists, ML engineers and platform teams to ensure high-quality, reliable data flows across cloud environments, balancing data engineering with ML data needs.

Qualifications

  • 4+ years of experience in Data Engineering.
  • Strong Python and SQL skills.
  • Experience building data pipelines for ML or analytics systems.
  • Hands-on experience with Spark, Databricks, or similar distributed processing frameworks.
  • Experience with orchestration tools (Airflow or similar).
  • Experience in AWS, Azure, or GCP environments.
  • Familiarity with data quality validation and monitoring frameworks.
  • Understanding of feature engineering and model data lifecycle.

Responsibilities

  • Design and build scalable data pipelines for ML workflows.
  • Develop feature engineering and data preparation processes.
  • Implement batch and real-time data ingestion systems.
  • Ensure data quality, validation, and monitoring.
  • Collaborate with ML engineers to support model training and deployment.
  • Integrate pipelines with orchestration tools (Airflow or similar).
  • Optimize pipeline performance and cloud cost efficiency.
  • Maintain documentation and version control of data workflows.

Skills

Python
SQL
Spark
Databricks
Airflow
Feature engineering
Data Pipelines
Data Quality
Great Expectations
Kafka
AWS
Azure
GCP

Tools

Databricks
Kafka

Job description

  • Work Experience Python, SQL, Spark, Databricks, Airflow, Feature Engineering, Data Pipelines, Data Quality, Great Expectations, AWS, Azure, GCP, Kafka
  • Required Skills
    • Airflow
    • AWS
    • +20
  • Remote Job
Job Description

This is a remote position.

We are seeking an AI Data Engineer to design and build production‑grade data pipelines that power machine learning systems. This role focuses on creating scalable ingestion, transformation, and feature engineering workflows that support model training, evaluation, and real‑time inference.

You will work closely with Data Scientists, Machine Learning Engineers, and Platform teams to ensure high‑quality, reliable, and efficient data flows across cloud environments. The ideal candidate understands both traditional data engineering and the unique data needs of ML systems.

Key Responsibilities
  • Design and build scalable data pipelines for ML workflows
  • Develop feature engineering and data preparation processes
  • Implement batch and real‑time data ingestion systems
  • Ensure data quality, validation, and monitoring
  • Collaborate with ML engineers to support model training and deployment
  • Integrate pipelines with orchestration tools (Airflow or similar)
  • Optimize pipeline performance and cloud cost efficiency
  • Maintain documentation and version control of data workflows
Requirements
  • 4+ years of experience in Data Engineering
  • Strong Python and SQL skills
  • Experience building data pipelines for ML or analytics systems
  • Hands‑on experience with Spark, Databricks, or similar distributed processing frameworks
  • Experience with orchestration tools (Airflow or similar)
  • Experience in AWS, Azure, or GCP environments
  • Familiarity with data quality validation and monitoring frameworks
  • Understanding of feature engineering and model data lifecycle
Preferred Qualifications
  • Experience with streaming systems (Kafka, Kinesis, Pub/Sub)
  • Experience supporting model deployment and MLOps workflows
  • Experience with feature stores or vector databases
  • Familiarity with ML frameworks (TensorFlow, PyTorch)
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Data Engineer (Data Engineering, Cloud Platform, Python)
AI Data Engineer (Data Engineering, Cloud Platform, Python)

Capgemini • Warszawa

On-site
PLN 120,000 - 160,000
Medical care
Insurance
Wellness resources
+1
Data & AI Engineer
Data & AI Engineer

Avanade • Kraków

On-site
PLN 180,000 - 320,000
Applied Machine Learning Engineer
Applied Machine Learning Engineer

Hitachi Energy • Kraków

On-site
PLN 45,000 - 65,000
Competitive benefits for financial wellbeing
Support for physical and mental wellbeing
Remote Senior Data Engineer — Databricks Pipelines & AI
Remote Senior Data Engineer — Databricks Pipelines & AI

Sowelo Consulting sp. z o.o. • Poland

On-site
PLN 180,000 - 240,000
Playfully remote-friendly
Remote AI Data Engineer: ML Pipelines & Feature Engineering
Remote AI Data Engineer: ML Pipelines & Feature Engineering

Zohorecruit • Polska

Hybrid
PLN 190,000 - 260,000
Remote work
Databricks Data Engineer
Databricks Data Engineer

RemoDevs • Warszawa, Kraków

On-site
PLN 180,000 - 280,000
Azure Data Engineering Architect - freelance
Azure Data Engineering Architect - freelance

Lingarogroup • Poland

On-site
PLN 180,000 - 240,000
Cloud Data Engineer (Snowflake/Databricks) – Remote
Cloud Data Engineer (Snowflake/Databricks) – Remote

Zohorecruit • Polska

Hybrid
PLN 180,000 - 280,000
Senior Data Engineer (PST Overlap)
Senior Data Engineer (PST Overlap)

Appliscale • Kraków

On-site
PLN 180,000 - 260,000
AI Solutions Architect
AI Solutions Architect

Zohorecruit • Polska

Hybrid
PLN 469,000 - 704,000