Senior Data Engineer (Databricks Migration)

Sigma Software

Województwo małopolskie

On-site

PLN 180,000 - 240,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Sigma Software’s Data Engineering Center of Excellence is seeking a Senior Data Engineer to build scalable, high-performance data platforms for enterprise retail analytics. You will drive the migration of a large-scale platform from BigQuery to Databricks, design Lakehouse architectures with Delta Lake, and collaborate with international teams to deliver reliable, well-governed data solutions.

The role emphasizes CI/CD, cross-functional collaboration, and modernization of analytics workloads

Qualifications

  • 5+ years of professional experience as a Data Engineer.
  • Strong programming skills in Python and advanced SQL.
  • Hands-on commercial experience with Databricks.
  • Strong knowledge of Apache Spark, primarily PySpark.
  • Experience designing and building modern cloud-based data platforms.
  • Experience developing ETL / ELT pipelines and large-scale data processing solutions.
  • Hands-on experience with Delta Lake.
  • Experience with Spark Declarative Pipelines.
  • Experience with cluster monitoring, metrics analysis, and performance optimization.
  • Strong understanding of distributed data processing architectures.
  • Solid understanding of data warehousing concepts and dimensional modeling.
  • Experience with Airflow or similar orchestration tools.
  • Experience optimizing complex analytical SQL workloads.
  • Experience implementing CI / CD practices for data engineering platforms.
  • Strong troubleshooting and performance optimization skills.
  • Ability to work collaboratively in cross-functional international teams.
  • Upper-Intermediate or higher English level.

Responsibilities

  • Participate in migration of large-scale analytics platform from BigQuery to Databricks.
  • Design and implement scalable Lakehouse architectures using Databricks and Delta Lake.
  • Analyze existing ETL/ELT workloads and define migration approaches.
  • Develop and optimize data pipelines processing large volumes of retail and analytical data.
  • Implement incremental processing strategies and scalable transformation frameworks.
  • Build and maintain Spark-based data processing solutions using PySpark.
  • Design and maintain medallion architecture layers Bronze, Silver, and Gold.
  • Implement data governance and security best practices using Unity Catalog.
  • Collaborate with Data Science, Analytics, Product, and Customer Engineering teams.
  • Participate in architecture discussions and technical solution design.
  • Develop reusable data platform components and engineering standards.
  • Conduct code reviews and contribute to platform reliability and maintainability.
  • Troubleshoot and optimize complex SQL and Spark workloads.
  • Support production deployments and platform modernization activities.

Skills

Python
SQL
Databricks
PySpark
Apache Spark
Delta Lake
Airflow
CI/CD
Unity Catalog
Cloud Data Platform
BigQuery

Tools

Databricks
Apache Spark
PySpark
Delta Lake
Python
SQL
Airflow
GCP
CI/CD
Unity Catalog

Job description

Are you a Senior Data Engineer passionate about building scalable, high-performance data platforms and working with modern Lakehouse technologies? Join Sigma Software’s Data Engineering Center of Excellence and contribute to the modernization of an enterprise-scale analytics ecosystem for the retail domain.
We are looking for a Senior specialist with strong Databricks, PySpark, and cloud data engineering expertise to participate in the migration of a large-scale analytical platform from BigQuery to Databricks. You will collaborate with international teams, contribute to architectural decisions, and help shape reliable and scalable data solutions.
We at Sigma Software create opportunities for continuous learning, technology growth, and meaningful engineering impact while working on complex international projects.

CUSTOMER

Our Customer is a leading retail technology company specializing in AI-driven pricing optimization solutions for enterprise retailers. The company helps businesses improve profitability and competitiveness through advanced analytics, automation, and intelligent pricing strategies. Their platform combines business intelligence with sophisticated algorithms to support data-informed pricing decisions at scale for global retail organizations.

PROJECT

The project focuses on the strategic migration of a large-scale analytical platform from a legacy BigQuery ecosystem to a modern Databricks Lakehouse architecture. The platform processes high-volume retail datasets, machine learning workloads, analytics pipelines, and customer-specific business logic.
As part of the modernization initiative, the engineering team is implementing scalable Spark-based processing, Delta Lake architecture, medallion data layers, and modern governance practices. The role offers an opportunity to work with distributed data processing systems, optimize large-scale workloads, and contribute to the evolution of an enterprise-grade data platform.

Key Technologies: Databricks, Apache Spark, PySpark, Delta Lake, Python, SQL, Airflow, GCP, CI/CD, Unity Catalog

  • Participate in the migration of a large-scale analytical platform from BigQuery to Databricks
  • Design and implement scalable Lakehouse architectures using Databricks and Delta Lake
  • Analyze existing ETL / ELT workloads and define migration approaches
  • Develop and optimize data pipelines processing large volumes of retail and analytical data
  • Implement incremental processing strategies and scalable transformation frameworks
  • Build and maintain Spark-based data processing solutions using PySpark
  • Design and maintain medallion architecture layers including Bronze, Silver, and Gold
  • Implement data governance and security best practices using Unity Catalog
  • Collaborate with Data Science, Analytics, Product, and Customer Engineering teams
  • Participate in architecture discussions and technical solution design
  • Develop reusable data platform components and engineering standards
  • Conduct code reviews and contribute to platform reliability and maintainability
  • Troubleshoot and optimize complex SQL and Spark workloads
  • Support production deployments and platform modernization activities
  • 5+ years of professional experience as a Data Engineer
  • Strong programming skills in Python and advanced SQL
  • Hands-on commercial experience with Databricks
  • Strong knowledge of Apache Spark, primarily PySpark
  • Experience designing and building modern cloud-based data platforms
  • Experience developing ETL / ELT pipelines and large-scale data processing solutions
  • Hands-on experience with Delta Lake
  • Experience with Spark Declarative Pipelines
  • Experience with cluster monitoring, metrics analysis, and performance optimization
  • Strong understanding of distributed data processing architectures
  • Solid understanding of data warehousing concepts and dimensional modeling
  • Experience with Airflow or similar orchestration tools
  • Experience optimizing complex analytical SQL workloads
  • Experience implementing CI / CD practices for data engineering platforms
  • Strong troubleshooting and performance optimization skills
  • Ability to work collaboratively in cross-functional international teams
  • Upper-Intermediate or higher English level
WILL BE A PLUS
  • Experience working with GCP cloud services
  • Experience with AWS or Azure cloud platforms
  • Experience in retail analytics or pricing optimization domains
  • Experience supporting machine learning or AI-related data workloads
  • Experience with platform modernization and cloud migration initiatives
PERSONAL PROFILE
  • Strong analytical and problem-solving mindset
  • Proactive and ownership-driven approach
  • Ability to work independently and collaboratively
  • Good communication and stakeholder collaboration skills
  • Passion for scalable data engineering and modern data platforms
  • Interest in continuous learning and technology innovation
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer (Databricks Migration)
Senior Data Engineer (Databricks Migration)

Sigma Software • Województwo mazowieckie

Hybrid
PLN 240,000 - 360,000
Senior Data Engineer: Databricks Lakehouse Migration
Senior Data Engineer: Databricks Lakehouse Migration

Sigma Software • Województwo małopolskie

On-site
PLN 180,000 - 240,000
Senior Data Engineer
Senior Data Engineer

Sigma Software • Warszawa

On-site
PLN 240,000 - 360,000
Senior Big Data Engineer (Databricks + Azure)
Senior Big Data Engineer (Databricks + Azure)

SoftServe • Poland

On-site
PLN 180,000 - 320,000
Data Platform Engineer
Data Platform Engineer

Sigma Embedded Engineering • Kraków

Hybrid
Long-term, stable projects
Work with modern technologies
Competitive salary and benefits
+2
Senior Data Engineer (Azure/Databricks)
Senior Data Engineer (Azure/Databricks)

HeadHR • Poznań

On-site
PLN 120,000 - 150,000
Private medical care
Co-financing for the sports card
Constant support of dedicated consultant
+1
data engineer for cloud-native data platforms
data engineer for cloud-native data platforms

Enfint • Warszawa

On-site
PLN 180,000 - 300,000
Senior Data Engineer (M/F/D)
Senior Data Engineer (M/F/D)

DSV Road GmbH • Warszawa

On-site
PLN 180,000 - 260,000
Private medical care
Parking space for employees
Life insurance
+2
Big Data Architect (Databricks + Azure)
Big Data Architect (Databricks + Azure)

SoftServe • Poland

On-site
Career development plans
Access to Udemy courses
Support for professional certifications
Senior Data Engineer
Senior Data Engineer

United States Digital Space LLC • Poland

Hybrid
PLN 180,000 - 300,000
International projects
In-office, hybrid, or remote flex
Medical healthcare
+7