Senior Data Engineer (Databricks Migration)

Sigma Software

Województwo mazowieckie

Hybrid

PLN 240,000 - 360,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Sigma Software is seeking a Senior Data Engineer to join its Data Engineering Center of Excellence in Warsaw/Województwo mazowieckie. You will lead migration from BigQuery to Databricks, design scalable Lakehouse architectures, and optimize large-scale data pipelines for a retail analytics platform.

Ideal candidates have 5+ years of data engineering, strong Python/SQL, Databricks, Spark, and Delta Lake experience, plus cloud expertise and CI/CD practices.

Qualifications

  • 5+ years of professional experience as a Data Engineer.
  • Strong programming skills in Python and advanced SQL.
  • Hands-on commercial experience with Databricks.
  • Strong knowledge of Apache Spark, primarily PySpark.
  • Experience designing and building modern cloud-based data platforms.
  • Experience developing ETL / ELT pipelines and large-scale data processing solutions.
  • Hands-on experience with Delta Lake.
  • Experience with Spark Declarative Pipelines.
  • Experience with cluster monitoring, metrics analysis, and performance optimization.
  • Strong understanding of distributed data processing architectures.
  • Solid understanding of data warehousing concepts and dimensional modeling.
  • Experience with Airflow or similar orchestration tools.
  • Experience optimizing complex analytical SQL workloads.
  • Experience implementing CI / CD practices for data engineering platforms.
  • Strong troubleshooting and performance optimization skills.
  • Ability to work collaboratively in cross-functional international teams.
  • Upper-Intermediate or higher English level.

Responsibilities

  • Participate in migration of large-scale analytical platform from BigQuery to Databricks.
  • Design and implement scalable Lakehouse architectures using Databricks and Delta Lake.
  • Analyze ETL/ELT workloads and define migration approaches.
  • Develop and optimize data pipelines for large retail datasets and analytics tasks.
  • Implement incremental processing strategies and scalable transformation frameworks.
  • Build and maintain Spark-based data processing solutions with PySpark.
  • Design and maintain medallion architecture layers (Bronze, Silver, Gold).
  • Implement data governance and security best practices using Unity Catalog.
  • Collaborate with Data Science, Analytics, Product, and Customer Engineering teams.
  • Participate in architecture discussions and technical solution design.
  • Develop reusable data platform components and engineering standards.
  • Conduct code reviews and contribute to platform reliability and maintainability.
  • Troubleshoot and optimize complex SQL and Spark workloads.
  • Support production deployments and platform modernization activities.

Skills

Python
SQL
Cross-functional collaboration
English proficient

Tools

Databricks
Apache Spark
PySpark
Delta Lake
Airflow
CI/CD

Job description

Are you a Senior Data Engineer passionate about building scalable, high-performance data platforms and working with modern Lakehouse technologies? Join Sigma Software’s Data Engineering Center of Excellence and contribute to the modernization of an enterprise-scale analytics ecosystem for the retail domain.
We are looking for a Senior specialist with strong Databricks, PySpark, and cloud data engineering expertise to participate in the migration of a large-scale analytical platform from BigQuery to Databricks. You will collaborate with international teams, contribute to architectural decisions, and help shape reliable and scalable data solutions.
We at Sigma Software create opportunities for continuous learning, technology growth, and meaningful engineering impact while working on complex international projects.

CUSTOMER

Our Customer is a leading retail technology company specializing in AI-driven pricing optimization solutions for enterprise retailers. The company helps businesses improve profitability and competitiveness through advanced analytics, automation, and intelligent pricing strategies. Their platform combines business intelligence with sophisticated algorithms to support data-informed pricing decisions at scale for global retail organizations.

PROJECT

The project focuses on the strategic migration of a large-scale analytical platform from a legacy BigQuery ecosystem to a modern Databricks Lakehouse architecture. The platform processes high-volume retail datasets, machine learning workloads, analytics pipelines, and customer-specific business logic.
As part of the modernization initiative, the engineering team is implementing scalable Spark-based processing, Delta Lake architecture, medallion data layers, and modern governance practices. The role offers an opportunity to work with distributed data processing systems, optimize large-scale workloads, and contribute to the evolution of an enterprise-grade data platform.

Key Technologies: Databricks, Apache Spark, PySpark, Delta Lake, Python, SQL, Airflow, GCP, CI/CD, Unity Catalog

  • Participate in the migration of a large-scale analytical platform from BigQuery to Databricks
  • Design and implement scalable Lakehouse architectures using Databricks and Delta Lake
  • Analyze existing ETL / ELT workloads and define migration approaches
  • Develop and optimize data pipelines processing large volumes of retail and analytical data
  • Implement incremental processing strategies and scalable transformation frameworks
  • Build and maintain Spark-based data processing solutions using PySpark
  • Design and maintain medallion architecture layers including Bronze, Silver, and Gold
  • Implement data governance and security best practices using Unity Catalog
  • Collaborate with Data Science, Analytics, Product, and Customer Engineering teams
  • Participate in architecture discussions and technical solution design
  • Develop reusable data platform components and engineering standards
  • Conduct code reviews and contribute to platform reliability and maintainability
  • Troubleshoot and optimize complex SQL and Spark workloads
  • Support production deployments and platform modernization activities
  • 5+ years of professional experience as a Data Engineer
  • Strong programming skills in Python and advanced SQL
  • Hands-on commercial experience with Databricks
  • Strong knowledge of Apache Spark, primarily PySpark
  • Experience designing and building modern cloud-based data platforms
  • Experience developing ETL / ELT pipelines and large-scale data processing solutions
  • Hands-on experience with Delta Lake
  • Experience with Spark Declarative Pipelines
  • Experience with cluster monitoring, metrics analysis, and performance optimization
  • Strong understanding of distributed data processing architectures
  • Solid understanding of data warehousing concepts and dimensional modeling
  • Experience with Airflow or similar orchestration tools
  • Experience optimizing complex analytical SQL workloads
  • Experience implementing CI / CD practices for data engineering platforms
  • Strong troubleshooting and performance optimization skills
  • Ability to work collaboratively in cross-functional international teams
  • Upper-Intermediate or higher English level
WILL BE A PLUS
  • Experience working with GCP cloud services
  • Experience with AWS or Azure cloud platforms
  • Experience in retail analytics or pricing optimization domains
  • Experience supporting machine learning or AI-related data workloads
  • Experience with platform modernization and cloud migration initiatives
PERSONAL PROFILE
  • Strong analytical and problem-solving mindset
  • Proactive and ownership-driven approach
  • Ability to work independently and collaboratively
  • Good communication and stakeholder collaboration skills
  • Passion for scalable data engineering and modern data platforms
  • Interest in continuous learning and technology innovation
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer (Databricks Migration)
Senior Data Engineer (Databricks Migration)

Sigma Software • Województwo małopolskie

On-site
PLN 180,000 - 240,000
Senior Data Engineer: Databricks Lakehouse Migration
Senior Data Engineer: Databricks Lakehouse Migration

Sigma Software • Województwo małopolskie

On-site
PLN 180,000 - 240,000
Senior Data Engineer
Senior Data Engineer

Sigma Software • Warszawa

On-site
PLN 240,000 - 360,000
Senior Big Data Engineer (Databricks + Azure)
Senior Big Data Engineer (Databricks + Azure)

SoftServe • Poland

On-site
PLN 180,000 - 320,000
Data Platform Engineer
Data Platform Engineer

Sigma Embedded Engineering • Kraków

Hybrid
Long-term, stable projects
Work with modern technologies
Competitive salary and benefits
+2
Senior Data Engineer (Azure/Databricks)
Senior Data Engineer (Azure/Databricks)

HeadHR • Poznań

On-site
PLN 120,000 - 150,000
Private medical care
Co-financing for the sports card
Constant support of dedicated consultant
+1
data engineer for cloud-native data platforms
data engineer for cloud-native data platforms

Enfint • Warszawa

On-site
PLN 180,000 - 300,000
Senior Data Engineer (M/F/D)
Senior Data Engineer (M/F/D)

DSV Road GmbH • Warszawa

On-site
PLN 180,000 - 260,000
Private medical care
Parking space for employees
Life insurance
+2
Big Data Architect (Databricks + Azure)
Big Data Architect (Databricks + Azure)

SoftServe • Poland

On-site
Career development plans
Access to Udemy courses
Support for professional certifications
Senior Data Engineer - Databricks Lakehouse Migration
Senior Data Engineer - Databricks Lakehouse Migration

Sigma Software • Województwo mazowieckie

Hybrid
PLN 240,000 - 360,000