A data engineering company based in Utrecht is seeking an experienced professional to design and maintain ETL/ELT pipelines leveraging Azure Databricks and PySpark. The role involves working with various Azure services including Data Factory and SQL, implementing Delta Lake architecture, and deploying solutions using CI/CD pipelines. Strong collaborative skills are essential as you will be working with data architects and business stakeholders to deliver high-quality data solutions.
Responsibilities
Design, develop, and maintain ETL/ELT pipelines using Azure Databricks and PySpark.
Work with Azure services: ADF, ADLS, Azure SQL, Synapse, Key Vault.
Implement Delta Lake and manage data lake house architecture.
Collaborate with data architects, analysts, and business stakeholders.
Deploy solutions using CI/CD pipelines (Azure DevOps preferred).
Skills
Azure Databricks
PySpark
SQL
Data modeling
Data governance principles
devops
Terraform
Delta Lake
Tools
Azure Data Factory
Azure SQL
Synapse
ADLS Gen2
Azure DevOps
Unity Catalog
Job description
Responsibilities
Design, develop, and maintain ETL/ELT pipelines using Azure Databricks and PySpark
Work with Azure services: ADF, ADLS, Azure SQL, Synapse, Key Vault
Implement Delta Lake and manage data lake house architecture
Collaborate with data architects, analysts, and business stakeholders
Deploy solutions using CI/CD pipelines (Azure DevOps preferred)
Required Skills
Strong experience with Azure Databricks, PySpark, Spark SQL
Hands-on with Azure Data Factory, Delta Lake, ADLS Gen2
Proficient in SQL, data modeling, and data governance principles
Experience with devops and Terraform
Familiarity with Lakehouse architecture and Unity Catalog (bonus)