Senior Data Pipeline Engineer (Batch & Streaming)

amo

Paris

Sur place

EUR 70 000 - 110 000

Plein temps

14 jours+

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Avantages offerts par ce poste

Onsite 5 days a week
Central Paris office near Opera

Résumé du poste

amo is seeking a senior data engineer to own and operate batch and streaming data pipelines in a production environment in Paris. You will move data into archival storage, improve data quality and observability, and design validations for freshness, completeness, and schema correctness.

You will build reusable datasets and optimize queries for cost and latency. Expect strong Python/SQL skills, production experience with Python or Scala, and familiarity with PySpark, Spark SQL, Spark Streaming,

Qualifications

  • Senior-level experience building and operating data pipelines.
  • Strong Python and SQL.
  • Production experience with Python or Scala.
  • Experience with PySpark, Spark SQL, Spark Streaming, Parquet, and Iceberg.
  • Strong understanding of batch and stream processing.
  • Query optimization experience.
  • Familiarity with data engine internals and distributed data systems.
  • Experience with data lakes, table/storage formats, and derived datasets.

Responsabilités

  • Build and maintain batch and streaming data pipelines.
  • Move production data into archival/data lake storage safely and observably.
  • Improve quality, reliability, and observability of derived datasets.
  • Design validations for freshness, completeness, schema correctness, and data quality.
  • Build reusable datasets that reduce one-off data extraction and aggregation work.
  • Optimize queries and data processing jobs for correctness, latency, and cost.
  • Prototype in the right tools, then integrate scoped production experiments with amo’s systems where needed.

Connaissances

Python
MySQL/SQL
Scala
Distributed data systems
Batch and streaming processing
Query optimization
Data lake/storage
Data pipelines

Outils

PySpark
Spark SQL
Spark Streaming
Parquet
Iceberg

Description du poste

amo is seeking a senior data engineer to own and operate batch and streaming data pipelines in a production environment in Paris. You will move data into archival storage, improve data quality and observability, and design validations for freshness, completeness, and schema correctness.

You will build reusable datasets and optimize queries for cost and latency. Expect strong Python/SQL skills, production experience with Python or Scala, and familiarity with PySpark, Spark SQL, Spark Streaming,

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

Senior Data Pipeline Architect — Paris Onsite
Senior Data Pipeline Architect — Paris Onsite

amo • Paris

Sur place
EUR 75 000 - 115 000
Health care 100% coverage
Office in central Paris near Opera
Long vacation entitlement (8–9 weeks)
Senior Data Engineer
Senior Data Engineer

Hashlist • Poissy

Sur place
EUR 85 000 - 125 000
Senior Data Engineer: Real-Time Pipelines & Spark
Senior Data Engineer: Real-Time Pipelines & Spark

Hashlist • Poissy

Sur place
EUR 85 000 - 125 000
Data Engineer: Scalable Pipelines & AI-Driven Data
Data Engineer: Scalable Pipelines & AI-Driven Data

ManoMano • Paris

Sur place
EUR 55 000 - 75 000
Paris Data Engineer: Real-Time Pipelines & Trusted Data
Paris Data Engineer: Real-Time Pipelines & Trusted Data

The French Sourcer • Paris

Sur place
EUR 75 000 - 110 000
Health insurance
Meal vouchers
Remote-friendly policy
Senior Data Engineer: Build Scalable Data Pipelines & AI
Senior Data Engineer: Build Scalable Data Pipelines & AI

Technology & Strategy • Paris

Sur place
EUR 45 000 - 60 000
Senior Data Engineer - Scala/AWS for Scalable Pipelines
Senior Data Engineer - Scala/AWS for Scalable Pipelines

Opus Recruitment Solutions • Lyon

Sur place
EUR 60 000 - 90 000
Senior Data Engineer: AI-Powered Data Pipelines
Senior Data Engineer: AI-Powered Data Pipelines

Voodoo • Paris

Hybride
EUR 90 000 - 140 000
Competitive salary based on experience
Swile Lunch voucher
Gymlib (100% covered by Voodoo)
+3
Data Platform Engineer — PySpark, Lakehouse & Databricks
Data Platform Engineer — PySpark, Lakehouse & Databricks

Launchmetrics • Paris

Hybride
EUR 90 000 - 120 000
Learning & development allowance
Benefits package tailored to location
Flexible working arrangements
+1
Data Engineer
Data Engineer

Move2cloud • Marseille

Sur place
EUR 45 000 - 70 000