Data Engineer (Databricks)

Addepto

Polska

Hybrid

PLN 180,000 - 320,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Remote/hybrid work options
Career growth and training
Databricks and Anthropic partnership

Job summary

Addepto, a leading AI consulting and data engineering company, seeks a Senior Big Data Engineer to design and build scalable data platforms using Databricks, Spark, Airflow, and Dagster. You will work on global projects across aerospace, energy, and automotive sectors, guiding architecture and data governance.

With flexible remote/hybrid arrangements, you will collaborate with Data Science teams, advance CI/CD/MLOps practices, and leverage partnerships with Databricks and Anthropic for

Qualifications

  • 3+ years of commercial Big Data experience.
  • Proficient in Python with clean code and OOP.
  • Strong SQL skills including performance tuning and data warehousing experience.
  • Experience designing data governance and data management processes.
  • Deep expertise with Databricks, Spark, Airflow, and related tooling.
  • Experience deploying solutions in Azure cloud environments.
  • Ability to translate business needs into scalable data solutions.

Responsibilities

  • Design and optimize scalable data processing pipelines for streaming and batch workloads using Databricks, Airflow, and Dagster.
  • Architect end-to-end data platforms with high availability and reliability.
  • Lead CI/CD and MLOps for automated deployments and model lifecycle management.
  • Develop applications to aggregate, process, and analyze data from diverse sources for performance and scalability.
  • Collaborate with Data Science teams on ML projects, including feature engineering and deployment.
  • Manage data transformations using Databricks, DBT, and Airflow to ensure integrity.
  • Translate business requirements into scalable technical solutions while ensuring data quality and governance.

Skills

Python programming
SQL & data warehousing
Big Data technologies
Consulting experience
Independent work

Education

Master’s or Ph.D. in CS/DS/Math/Physics

Tools

Databricks
Spark
Apache Airflow
Dagster
Azure
Power BI

Job description

O projekcie:

Addepto is a leading AI consulting (addepto.com/ai-consulting/) and data engineering (addepto.com/data-engineering-services/) company that builds scalable, ROI-focused AI solutions for some of the world's largest enterprises and pioneering startups, including Rolls Royce, Continental, Porsche, ABB, and WGU. With an exclusive focus on Artificial Intelligence and Big Data, Addepto helps organizations unlock the full potential of their data through systems designed for measurable business impact and long-term growth.

  • Design and development of a universal data platform for global aerospace companies. This Azure and Databricks powered initiative combines diverse enterprise and public data sources. The data platform is at the early stages of the development, covering design of architecture and processes, as well as giving freedom for technology selection.
  • Data Platform Transformation for energy management association body. This project addressed critical data management challenges, boosting user adoption, performance, and data integrity. The team is implementing a comprehensive data catalog, leveraging Databricks and Apache Spark/PySpark, for simplified data access and governance. Secure integration solutions and enhanced data quality monitoring, utilizing Delta Live Table tests, established trust in the platform. The intermediate result is a user-friendly, secure, and data-driven platform, serving as a basis for further development of ML components.
  • Design of the data transformation and following data ops pipelines for global car manufacturer. This project aims to build a data processing system for both real-time streaming and batch data. We’ll handle data for business uses like process monitoring, analysis, and reporting, while also exploring LLMs for chatbots and data analysis. Key tasks include data cleaning, normalization, and optimizing the data model for performance and accuracy.
Discover our perks and benefits:
  • Work in a supportive team of passionate enthusiasts of AI & Big Data.
  • Engage with top-tier global enterprises and cutting-edge startups on international projects.
  • Enjoy flexible work arrangements, allowing you to work remotely or from modern offices and coworking spaces.
  • Accelerate your professional growth through career paths, knowledge-sharing initiatives, language classes, and sponsored training or conferences, including a partnership with Databricks and Anthropic, which offers industry-leading training materials and certifications.
  • Participate in team-building events and utilize the integration budget.
  • Celebrate work anniversaries, birthdays, and milestones.
  • Access medical and sports packages, eye care, and well-being support services, including psychotherapy and coaching.
  • Get full work equipment for optimal productivity, including a laptop and other necessary devices.
  • With our backing, you can boost your personal brand by speaking at conferences, writing for our blog, or participating in meetups.
Important:

Addepto may publish job opportunities on external job boards, but all applications are collected and processed through our official recruitment system: Recruitee addepto.recruitee.com

Please make sure that any further recruitment communication you receive comes from an official Addepto recruitment channel. If you receive a job offer, interview invitation, or recruitment email from an unofficial or unverified source, please be cautious, as it may not come from Addepto.

Wymagania:
  • At least 3 years of commercial experience implementing, developing, or maintaining Big Data systems.
  • Strong programming skills in Python: writing a clean code, OOP design.
  • Strong SQL skills, including performance tuning, query optimization, and experience with data warehousing solutions.
  • Experience in designing and implementing data governance and data management processes.
  • Deep expertise in Big Data technologies, including Databricks, Spark, Apache Airflow and other modern data orchestration and transformation tools.
  • Experience implementing and deploying solutions in cloud environments (with a preference for Azure).
  • Knowledge of how to build and deploy Power BI reports and dashboards for data visualization.
  • Excellent understanding of dimensional data and data modeling techniques.
  • Consulting experience and the ability to guide clients through architectural decisions, technology selection, and best practices.
  • Ability to work independently and take ownership of project deliverables.
  • Master’s or Ph.D. in Computer Science, Data Science, Mathematics, Physics, or a related field.
Codzienne zadania:
  • Design and optimize scalable data processing pipelines for both streaming and batch workloads using Big Data technologies such as Databricks, Apache Airflow, and Dagster.
  • Architect and implement end-to-end data platforms, ensuring high availability, performance, and reliability.
  • Lead the development of CI/CD and MLOps processes to automate deployments, monitoring, and model lifecycle management.
  • Develop and maintain applications for aggregating, processing, and analyzing data from diverse sources, ensuring efficiency and scalability.
  • Collaborate with Data Science teams on Machine Learning projects, including text/image analysis, feature engineering, and predictive model deployment.
  • Design and manage complex data transformations using Databricks, DBT, and Apache Airflow, ensuring data integrity and consistency.
  • Translate business requirements into scalable and efficient technical solutions while ensuring optimal performance and data quality.
  • Ensure data security, compliance, and governance best practices are followed across all data pipelines.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Data Engineer (Azure Data Factory)
Senior Data Engineer (Azure Data Factory)

Addepto • Polska

Hybrid
PLN 180,000 - 240,000
Remote work flexibility
Professional development & training
Medical and well-being packages
+1
Data Engineer
Data Engineer

Addepto • Warszawa

On-site
PLN 170,212 - 255,319
Fast career path and training sponsorship
Flexible working hours
Paid vacation
+1
Senior / Lead Data Software Engineer (Python, Spark, Azure)
Senior / Lead Data Software Engineer (Python, Spark, Azure)

EPAM Systems (Poland) sp. z o.o. • Wrocław

On-site
PLN 240,000 - 360,000
Ubezpieczenie zdrowotne
Tryb pracy hybrydowy
Praca zdalna na terenie Polski
+1
Lead Data Engineer (Databricks)
Lead Data Engineer (Databricks)

ACAISOFT POLAND Sp. z o.o. • Warszawa

On-site
PLN 180,000 - 280,000
Private medical care
Multisport card
Friendly, informal atmosphere, and in‑
Lead Data Scientist / ML Engineer
Lead Data Scientist / ML Engineer

Addepto • Poland

On-site
PLN 180,000 - 300,000
Remote work
Office/Co-working spaces
Training & Conferences
+2
Data consultant
Data consultant

AtlantisJobs • Polska

Remote
PLN 110,000 - 160,000
Expert Data Engineer
Expert Data Engineer

GFT Poland • Poland

Hybrid
PLN 200,000 - 320,000
Hybd (Lodz, Poznań, Kraków, Warszawa,W
Pakiet benefitów
Data Engineer (Spark)
Data Engineer (Spark)

Addepto • Województwo pomorskie

On-site
PLN 120,000 - 180,000
Flexible work arrangements
Remote or office options
Professional development opportunities
Data Engineer Consultant
Data Engineer Consultant

Chabre • Warszawa

On-site
PLN 198,000 - 331,000
Rate up to 240 PLN/h + VAT
Peripherals subsidy 500 PLN
Work tools provided
+1
Data Engineer (Spark)
Data Engineer (Spark)

Addepto • Białystok

Hybrid
PLN 180,000 - 210,000
Remote work opportunities
Flexible working hours
Training & conference budget
+2