Data Engineer Consultant

Chabre

Warszawa

On-site

PLN 198,000 - 331,000

Full time

9 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Rate up to 240 PLN/h + VAT
Peripherals subsidy 500 PLN
Work tools provided
Course reimbursement

Job summary

Chabre IT Services is seeking a Data Engineer Consultant in Warsaw to design and implement data ingestion pipelines on the Databricks platform, connecting external APIs and business systems for scalable data flows.

You will develop Bronze and Silver layers, implement data quality controls, and support ML workflows, including feature stores and batch scoring. Excellent English and Polish (min B2) are required for successful collaboration on client projects.

Qualifications

  • Recent, hands-on experience delivering production solutions with Databricks.
  • Strong knowledge of PySpark, Delta Lake and Unity Catalog.
  • Practical experience with the Databricks medallion architecture and production data pipelines.
  • Experience with Spark Declarative Pipelines / Delta Live Tables and Databricks Workflows.
  • Very good Python and advanced SQL skills.
  • Experience integrating REST APIs into data engineering solutions, including authentication, pagination, rate limiting and incremental or checkpointed data retrieval.
  • Strong understanding of data modelling, including SCD Type 2 and CDC patterns.
  • Experience designing data quality controls, reconciliation processes and mechanisms for handling invalid or quarantined data.
  • Knowledge of Git-based development and CI/CD for data platforms, preferably with Azure DevOps or GitHub Actions.
  • Experience implementing monitoring and observability for production data pipelines.
  • Ability to work within a fixed project scope, estimate work, meet milestones and take ownership of agreed deliverables.
  • Strong communication skills and the ability to work independently in a consulting environment.
  • Excellent English and Polish skills (min. B2 level).

Responsibilities

  • Design and implement data ingestion pipelines connecting external APIs and business systems with the Databricks data platform.
  • Develop reliable processes for retrieving and processing historical data, including mechanisms for rate limiting, checkpointing, retries and safe reprocessing.
  • Build and maintain the Bronze layer of the data platform, covering data standardisation, typing, historical tracking and document storage.
  • Develop processes for linking digital documents with corresponding business records while ensuring data consistency and idempotent processing.
  • Use Databricks capabilities to classify documents and extract structured information from PDFs and images, including confidence scoring and workflows for manual review.
  • Contribute to the development of the Silver layer using appropriate entity modelling, CDC and incremental processing patterns.
  • Implement deterministic entity resolution and source-priority logic across multiple data sources.
  • Apply data quality controls within the existing framework and ensure that quality results are properly stored, monitored and available for governance purposes.
  • Support the development of analytical and machine-learning data products, including feature and metric stores and batch scoring workflows.
  • Work with MLflow and related processes for model registration and promotion where required.
  • Improve automation, monitoring and reliability across data pipelines and contribute to CI/CD processes.
  • Collaborate with other technical stakeholders to deliver agreed project milestones within a defined consulting scope.

Skills

Databricks
PySpark/Delta
Unity Catalog
Delta Live Tables
Python/SQL
REST APIs
Data modelling (SCD2/CDC)
Data quality controls
CI/CD / Git
Monitoring/Observability
Project ownership
Communication (EN/PL)
English & Polish (B2)

Tools

Databricks
Git-based development
Azure DevOps/GitHub Actions
CI/CD pipelines

Job description

Working as a Data Engineer Consultant, you will:

Design and implement data ingestion pipelines connecting external APIs and business systems with the Databricks data platform.

Develop reliable processes for retrieving and processing historical data, including mechanisms for rate limiting, checkpointing, retries and safe reprocessing.

Build and maintain the Bronze layer of the data platform, covering data standardisation, typing, historical tracking and document storage.

Develop processes for linking digital documents with corresponding business records while ensuring data consistency and idempotent processing.

Use Databricks capabilities to classify documents and extract structured information from PDFs and images, including confidence scoring and workflows for manual review.

Contribute to the development of the Silver layer using appropriate entity modelling, CDC and incremental processing patterns.

Implement deterministic entity resolution and source-priority logic across multiple data sources.

Apply data quality controls within the existing framework and ensure that quality results are properly stored, monitored and available for governance purposes.

Support the development of analytical and machine-learning data products, including feature and metric stores and batch scoring workflows.

Work with MLflow and related processes for model registration and promotion where required.

Improve automation, monitoring and reliability across data pipelines and contribute to CI/CD processes.

Collaborate with other technical stakeholders to deliver agreed project milestones within a defined consulting scope.

Chabre IT Services is a global professional IT services provider, building long-lasting relationships with Enterprises. We specialize in the delivery of tailor‑made solutions, smart outsourcing, try&hire, and success fee services. We are a smart IT boutique with unique knowledge, which will deliver your ideas into reality.

About our Client

Our client is a global technology services organisation delivering data and digital transformation projects for enterprise customers. The project focuses on building and enhancing a modern data platform based on Databricks, with a strong emphasis on data engineering, governance, automation and analytics.

The successful candidate will join a project involving multiple data sources, API integrations, document processing and layered data architecture. The environment combines traditional data engineering with modern Databricks AI capabilities and machine‑learning workflows, providing an opportunity to work across several stages of the data lifecycle.

Qualifications
  • Recent, hands‑on experience delivering production solutions with Databricks.
  • Strong knowledge of PySpark, Delta Lake and Unity Catalog.
  • Practical experience with the Databricks medallion architecture and production data pipelines.
  • Experience with Spark Declarative Pipelines / Delta Live Tables and Databricks Workflows.
  • Very good Python and advanced SQL skills.
  • Experience integrating REST APIs into data engineering solutions, including authentication, pagination, rate limiting and incremental or checkpointed data retrieval.
  • Strong understanding of data modelling, including SCD Type 2 and current‑state / CDC patterns.
  • Experience designing data quality controls, reconciliation processes and mechanisms for handling invalid or quarantined data.
  • Knowledge of Git‑based development and CI/CD for data platforms, preferably with Azure DevOps or GitHub Actions.
  • Experience implementing monitoring and observability for production data pipelines.
  • Ability to work within a fixed project scope, estimate work, meet milestones and take ownership of agreed deliverables.
  • Strong communication skills and the ability to work independently in a consulting environment.
  • Excellent English and Polish skills (min. B2 level).
We offer
  • Rate up to 240,00 PLN/h + VAT
  • Subsidy for peripherals in the amount of 500,00zł
  • Working tool (MacBook Pro or Lenovo Legion 5)
  • Co‑financing of courses related to the position
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Engineer (Databricks) | Enterprise Energy Data Platform
Data Engineer (Databricks) | Enterprise Energy Data Platform

Polcode • Warszawa

Remote
PLN 152,000 - 207,000
Multisport card
Private medical care
Life insurance
+2
Senior Data Engineer | B2B Cooperation
Senior Data Engineer | B2B Cooperation

Profitroom • Poland

On-site
PLN 257,000 - 301,000
B2B cooperation
International projects
Flexible collaboration setup
Data Engineer
Data Engineer

Billennium • Wrocław

On-site
PLN 180,000 - 280,000
Sports card and wellbeing initiatives
Training platforms and certification
International projects growth
+1
Data Engineer (Databricks)
Data Engineer (Databricks)

Addepto • Polska

Hybrid
PLN 180,000 - 320,000
Remote/hybrid work options
Career growth and training
Databricks and Anthropic partnership
Data Engineer Scala/Spark
Data Engineer Scala/Spark

DAC.digital Group • Województwo pomorskie

On-site
PLN 262,000 - 357,000
Remote work option
On-site optional at Gdańsk office
data engineer for AI and ML projects
data engineer for AI and ML projects

Inetum • Polska

On-site
PLN 180,000 - 260,000
Funding training
Flexible hours
Cafeteria benefits
+7
Senior Data Engineer
Senior Data Engineer

MDW • Polska

Hybrid
PLN 180,000 - 260,000
25 days paid vacation
Training & certifications budget
Team building events
Databricks Engineer
Databricks Engineer

DAC.digital Group • Województwo pomorskie

On-site
PLN 167,000 - 279,000
Remote work
Data Architect (Databricks) (all genders)
Data Architect (Databricks) (all genders)

Lufthansa Systems • Województwo pomorskie

On-site
PLN 156,000 - 279,000
Multisport
Medical care
Life insurance
+1
Data Engineer (On-site) @ Capgemini Invent
Data Engineer (On-site) @ Capgemini Invent

Capgemini Invent • Województwo pomorskie

Hybrid
PLN 120,000 - 160,000
Training budget
Sport subscription
Private healthcare
+8