Databricks Lakehouse Data Engineer - Pipelines & Governance

Jobtailor

Lisboa

Presencial

EUR 50 000 - 70 000

Tempo integral

Há 4 dias
Torna-te num dos primeiros candidatos
Gerador de candidaturas

Transforma esta função numa entrevista — um currículo e uma carta de apresentação criados à volta do que este empregador procura.

Ultrapassa os filtros ATS

Resumo da oferta

Jobtailor in Lisbon is seeking a data engineer to build and maintain scalable data pipelines on the Databricks Lakehouse Platform using Spark and PySpark. You will ingest data from databases, files and APIs, model bronze/silver/gold transforms, and ensure data quality and governance.

You will optimize performance and costs, apply security and Unity Catalog governance, contribute to CI/CD for data assets, and collaborate with data scientists and analysts to deliver reliable datasets for analytics

Qualificações

  • Bachelor's or Master's degree in CS, IS, Engineering, or related field.
  • 2+ years of professional experience in data engineering or software engineering with a strong data component.
  • Solid Python skills in production environments.
  • Hands-on experience with Spark/PySpark for large-scale batch or streaming data processing.

Responsabilidades

  • Build and maintain scalable data processing pipelines on the Databricks Lakehouse Platform using Apache Spark and PySpark.
  • Ingest and integrate datasets from databases, file systems, and APIs into the Lakehouse using Auto Loader, CDC, and batch or streaming ingestion patterns.
  • Model and implement bronze, silver, and gold transformation layers using Delta Lake and Databricks SQL.
  • Ensure data integrity, consistency, and quality through validation and monitoring using Delta Live Tables expectations and Lakehouse monitoring capabilities.
  • Tune pipelines for performance and cost through Spark job optimization, Delta table maintenance, and cluster and compute configuration.
  • Apply data security, access control, and privacy practices using Unity Catalog for governance, permissions, and lineage.
  • Contribute to CI/CD and deployment automation for data assets.
  • Support migration initiatives from on-premise Cloudera environments to the Databricks Lakehouse.
  • Collaborate with data scientists and analysts to deliver reliable datasets for analytics and machine learning workflows.
  • Occasionally support lightweight Python services and APIs exposing data to downstream consumers.

Conhecimentos

Python
SQL
Apache Spark
PySpark
Data Modeling
Delta Lake
Delta Live Tables
Structured Streaming
Data Ingestion
CI/CD
Testing
Version Control
Collaboration
Communication

Formação académica

Bachelor's or Master's degree in Computer Science, Information Systems, Engineering, or a related field

Ferramentas

Databricks
Airflow
Kafka
Terraform

Descrição da oferta de emprego

Jobtailor in Lisbon is seeking a data engineer to build and maintain scalable data pipelines on the Databricks Lakehouse Platform using Spark and PySpark. You will ingest data from databases, files and APIs, model bronze/silver/gold transforms, and ensure data quality and governance.

You will optimize performance and costs, apply security and Unity Catalog governance, contribute to CI/CD for data assets, and collaborate with data scientists and analysts to deliver reliable datasets for analytics

Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Databricks Data Engineer - Lakehouse & Pipelines
Databricks Data Engineer - Lakehouse & Pipelines

Capgemini • Lisboa

Híbrido
EUR 35 000 - 50 000
Health and Life insurance
Career Acceleration Programs
Referral program with bonuses
Databricks Data Engineer — Pipelines & Governance
Databricks Data Engineer — Pipelines & Governance

SysMatch • Lisboa

Presencial
EUR 30 000 - 40 000
Remote Senior Data Engineer: Databricks Lakehouse Pipelines
Remote Senior Data Engineer: Databricks Lakehouse Pipelines

AgileEngine • Lisboa

Presencial
EUR 103 000 - 155 000
Growth opportunities
Competitive compensation
Remote-friendly
Senior Databricks Engineer – Lakehouse Data Pipelines
Senior Databricks Engineer – Lakehouse Data Pipelines

Cronos Europa • Portugal

Híbrido
EUR 70 000 - 110 000
Salary package
Work-life balance
Large-scale international platforms
+2
Senior Data Engineer - Databricks Lakehouse (Remote)
Senior Data Engineer - Databricks Lakehouse (Remote)

AgileEngine • Porto

Presencial
EUR 65 000 - 90 000
Growth without limits
Competitive compensation
Flexibility: fully remote
Senior Databricks Data Engineer – Lakehouse & CI/CD Expert
Senior Databricks Data Engineer – Lakehouse & CI/CD Expert

Capgemini • Lisboa

Híbrido
EUR 45 000 - 65 000
Health insurance
Life insurance
Referral bonuses
Senior Data Engineer — Remote Lakehouse & AI Pipelines
Senior Data Engineer — Remote Lakehouse & AI Pipelines

AgileEngine • Aveiro

Presencial
EUR 80 000 - 120 000
Growth opportunities
Competitive compensation
Remote-friendly with flexible hours
+3
Azure Data Engineer: Scalable Pipelines & Governance
Azure Data Engineer: Scalable Pipelines & Governance

JobCubby • Lisboa

Híbrido
EUR 45 000 - 65 000
Data Engineer: Build Real‑Time Pipelines & Lakehouse
Data Engineer: Build Real‑Time Pipelines & Lakehouse

KWAN • Lisboa

Presencial
EUR 45 000 - 75 000
Competitive salary
Awesome benefits
Career Growth Support
+1
Senior Databricks Data Engineer - Build Scalable Pipelines
Senior Databricks Data Engineer - Build Scalable Pipelines

EPAM Systems • Portugal

Presencial
EUR 60 000 - 90 000
Competitive compensation
Flexible work hours
Career growth opportunities
+1