Data Engineer — Ingestion & Lakehouse

Preply

Barcelona

Presencial

EUR 65.000 - 90.000

Jornada completa

14 días+

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Ventajas ofrecidas por este puesto de trabajo

Relocation package to Barcelona Hub
Health insurance
Learning & Development budget
Monthly lessons allowance

Descripción de la vacante

Preply is seeking a Data Engineer to join the Data Ingestion and Enrichment team. You will build and maintain the data lake components and production-grade ingestion pipelines powering analytics, ML, and product features.

You will collaborate with ML Platform, Analytics, and Product squads to ensure pipelines are observable, scalable, and reusable. You will implement data contracts, governance-by-design, and enrichment logic across domains, with a focus on reliability, lineage, and data quality.

Formación

  • Hands-on experience building components of large, high-scale applications (data pipelines, APIs, algorithms).
  • Experience in platform or data engineering teams in multi-stakeholder environments.
  • Cloud platforms (AWS/GCP or equivalent) and modern DevOps practices.
  • Hands-on design/implementation of real-time and batch data processing pipelines (Spark, Flink, Spark Streaming, Kafka, Debezium).
  • Experience with orchestration tools such as Airflow or dbt.
  • Strong problem-solving and proactive mindset with continuous improvement focus.
  • Excellent communication and cross-functional collaboration skills (English level B2+)

Responsabilidades

  • Contribute to trusted ingestion & enrichment foundations (Data Lake and Data as a Product).
  • Develop end-to-end ingestion pipelines (batch & streaming).
  • Define raw -> standardized -> consumption layers with lineage and retention strategies.
  • Implement data contracts, validation, anomaly detection, and quality checks.
  • Build enrichment logic and support versioning for downstream analytics.
  • Instrument ingestion pipelines with observability: freshness, latency, quality, cost.
  • Contribute to governance: access control, privacy protections, auditability.
  • Enable self-service through standardized templates and documentation.
  • Collaborate with Product, Backend, Analytics, and ML teams; mentor juniors.

Conocimientos

Problem-solving
Cross-functional collaboration
Proactive mindset
English B2+
Communication skills

Herramientas

Spark
Flink
Kafka
Debezium
Airflow
dbt

Descripción del empleo

Preply is seeking a Data Engineer to join the Data Ingestion and Enrichment team. You will build and maintain the data lake components and production-grade ingestion pipelines powering analytics, ML, and product features.

You will collaborate with ML Platform, Analytics, and Product squads to ensure pipelines are observable, scalable, and reusable. You will implement data contracts, governance-by-design, and enrichment logic across domains, with a focus on reliability, lineage, and data quality.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Data Engineer – Ingestion & Lakehouse Pipelines
Data Engineer – Ingestion & Lakehouse Pipelines

Preply • Bellprat

Presencial
EUR 60.000 - 90.000
Relocation package to Barcelona Hub
Health insurance
Learning & Development budget
+3
Senior Data Engineer — Ingestion & Lakehouse
Senior Data Engineer — Ingestion & Lakehouse

Preply • Barcelona

Presencial
EUR 90.000 - 130.000
Lessons allowance for Preply.com
Learning & Development budget
Health insurance
+3
Senior Data Engineer – Ingestion & Enrichment (Data Lakehouse)
Senior Data Engineer – Ingestion & Enrichment (Data Lakehouse)

Preply • Barcelona

Presencial
EUR 50.000 - 70.000
Data Engineer: Build Scalable Data Lakehouse & Pipelines
Data Engineer: Build Scalable Data Lakehouse & Pipelines

Vorwerk Austria GmbH & Co KG • Madrid

Presencial
EUR 50.000 - 78.000
Career growth
Training budget
Flexible hours
+8
Remote-friendly Data Platform Engineer - Lakehouse & PySpark
Remote-friendly Data Platform Engineer - Lakehouse & PySpark

PARKLU by Launchmetrics • Madrid

Presencial
EUR 60.000 - 90.000
Senior Data Engineer — Build a Global Lakehouse Platform
Senior Data Engineer — Build a Global Lakehouse Platform

Progress Partners • España

Presencial
EUR 70.000 - 100.000
Remote-first
Flexible hours
Unlimited PTO
+4
Senior Data Engineer - Cloud Data Lake & Streaming
Senior Data Engineer - Cloud Data Lake & Streaming

Parser • Madrid

Presencial
EUR 65.000 - 105.000
Knowledge sharing culture
Senior Data Engineer - Data Platform & Lakehouse Architect
Senior Data Engineer - Data Platform & Lakehouse Architect

Quantum Software • Donostia/San Sebastián

Híbrido
EUR 70.000 - 110.000
Indefinite contract
Equal pay guaranteed
Variable performance bonus
+7
Data Lakehouse Engineer for AI & Analytics
Data Lakehouse Engineer for AI & Analytics

Vorwerk Group • Madrid

Híbrido
EUR 40.000 - 60.000
25 days of vacation
Annual Variable Bonus
Private health insurance with Sanitas
+3
Senior Data Engineer: Lakehouse & Real-Time Pipelines
Senior Data Engineer: Lakehouse & Real-Time Pipelines

Capitole • España

Presencial
EUR 65.000 - 100.000
Formación 1.200€ al año
Retribución flexible en transporte y**
Eventos y team buildings
+2