Data Platform Engineer - PySpark & Lakehouse

Launchmetrics

Gerona

Presencial

EUR 65.000 - 90.000

Jornada completa

hace 4 horas
Sé de los primeros/as/es en solicitar esta vacante

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Ventajas ofrecidas por este puesto de trabajo

Learning & development allowance
Remote-friendly with hubs
Flexible working arrangements

Descripción de la vacante

Launchmetrics is seeking a Data Platform Engineer to design and build batch and near-real-time data pipelines using PySpark and Databricks. You’ll shape Delta Lake schemas and collaborate across product, QA, and data engineering to deliver scalable, reliable data enrichment for our media intelligence platform.

You’ll own code quality with unit testing, optimize pipelines for cost and performance, and contribute to cross‑pod initiatives in a fast-paced SaaS environment.

Formación

  • Engineer Degree or Master Degree in Computer Science and 3+ years of relevant work experience in full‑stack development in a SaaS environment
  • Strong Python and PySpark experience; comfort with distributed data processing at scale
  • Experience with a Lakehouse architecture (Databricks, Delta Lake, or comparable)
  • Familiarity with medallion architecture patterns (Bronze/Silver/Gold) or similar layered data design
  • Ability to reason about schema evolution, partitioning/clustering strategy, and pipeline reliability

Responsabilidades

  • Design and build data pipelines (batch and near‑real‑time) using PySpark and Databricks, across the Bronze/Silver/Gold medallion layers
  • Architect efficient Delta Lake table schemas — partitioning/liquid clustering strategy, schema evolution handling, and enrichment workflows
  • Work closely with product, QA, and other data engineers to translate enrichment and search requirements into reliable pipelines
  • Own code quality: structured PySpark jobs, unit tests (pytest), and adherence to team conventions
  • Continuously improve pipeline reliability and cost efficiency (OPTIMIZE scheduling, retry/backoff logic, concurrency handling)
  • Participate in cross‑pod initiatives across the data platform

Conocimientos

Python
PySpark
Distributed processing
Data pipelines
Batch processing
ETL
SQL

Educación

Bachelor/Master in Computer Science

Herramientas

Databricks
Delta Lake
Jira
GitHub
Serverless

Descripción del empleo

Launchmetrics is seeking a Data Platform Engineer to design and build batch and near-real-time data pipelines using PySpark and Databricks. You’ll shape Delta Lake schemas and collaborate across product, QA, and data engineering to deliver scalable, reliable data enrichment for our media intelligence platform.

You’ll own code quality with unit testing, optimize pipelines for cost and performance, and contribute to cross‑pod initiatives in a fast-paced SaaS environment.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Data Platform Engineer — Lakehouse & PySpark (Remote)
Data Platform Engineer — Lakehouse & PySpark (Remote)

Launchmetrics • Madrid

Híbrido
EUR 55.000 - 85.000
Remote-friendly Data Platform Engineer - Lakehouse & PySpark
Remote-friendly Data Platform Engineer - Lakehouse & PySpark

PARKLU by Launchmetrics • Madrid

Presencial
EUR 60.000 - 90.000
Hybrid Data Platform Engineer — PySpark & Delta Lake
Hybrid Data Platform Engineer — PySpark & Delta Lake

Launchmetrics • Barcelona

Híbrido
EUR 70.000 - 110.000
Data Platform Engineer — Remote-Friendly, Growth & Impact
Data Platform Engineer — Remote-Friendly, Growth & Impact

Launchmetrics • Gerona

Presencial
EUR 60.000 - 90.000
Learning and development allowance
Flexible working arrangements
Home office support
+1
Data Engineer
Data Engineer

Launchmetrics • Barcelona

Híbrido
EUR 70.000 - 110.000
Data Engineer at Launchmetrics
Data Engineer at Launchmetrics

Launchmetrics • Gerona

Presencial
EUR 60.000 - 90.000
Learning and development allowance
Flexible working arrangements
Home office support
+1
Data Engineer
Data Engineer

Launchmetrics • Madrid

Híbrido
EUR 55.000 - 85.000
Data Engineer – Ingestion & Lakehouse Pipelines
Data Engineer – Ingestion & Lakehouse Pipelines

Preply • Bellprat

Presencial
EUR 60.000 - 90.000
Relocation package to Barcelona Hub
Health insurance
Learning & Development budget
+3
Senior Spark Data Engineer - AWS & Data Pipelines
Senior Spark Data Engineer - AWS & Data Pipelines

Banco Santander SA • Madrid

Presencial
EUR 90.000 - 120.000
Data Engineer — Ingestion & Lakehouse
Data Engineer — Ingestion & Lakehouse

Preply • Barcelona

Presencial
EUR 65.000 - 90.000
Relocation package to Barcelona Hub
Health insurance
Learning & Development budget
+1