Senior Data Pipeline Architect — Paris Onsite

amo

Paris

Sur place

EUR 75 000 - 115 000

Plein temps

14 jours+

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Avantages offerts par ce poste

Health care 100% coverage
Office in central Paris near Opera
Long vacation entitlement (8–9 weeks)

Résumé du poste

amo is seeking a senior data engineer to build and operate robust data pipelines in Paris. The role focuses on batch and streaming processing, data lake storage, and producing high-quality, observable datasets.

You will optimize queries, validate data quality, and integrate production experiments within amo’s systems. The ideal candidate has strong Python/SQL skills, production experience with Python or Scala, and hands-on experience with PySpark, Spark SQL, Spark Streaming, Parquet and Iceberg.

Qualifications

  • Senior experience building and operating data pipelines.
  • Proficient in Python and SQL.
  • Production experience with Python or Scala.
  • Experience with PySpark, Spark SQL, Spark Streaming, Parquet, and Iceberg.
  • Strong understanding of batch and stream processing.
  • Experience with query optimization.
  • Familiarity with data engine internals and distributed data systems.
  • Experience with data lakes, table/storage formats, and derived datasets.

Responsabilités

  • Build and maintain batch and streaming data pipelines.
  • Move production data into archival/data lake storage safely and observably.
  • Improve quality, reliability, and observability of derived datasets.
  • Design validations for freshness, completeness, schema correctness, and data quality.
  • Build reusable datasets to reduce one-off data extraction and aggregation work.
  • Optimize queries and data processing jobs for correctness, latency, and cost.
  • Prototype in the right tools, then integrate scoped production experiments with amo’s systems where needed.

Connaissances

Senior data pipeline engineering
Python
SQL
Production experience with Python/Scal
Batch and stream processing
Query optimization
Distributed data systems
Data lakes & storage formats
Flink or similar stream processing
ML training/inference pipelines
Ranking/recommendations/embeddings
GDPR workflows / data privacy
Rust

Formation

Bachelor's degree preferred

Outils

PySpark
Spark SQL
Spark Streaming
Parquet
Iceberg
Scala

Description du poste

amo is seeking a senior data engineer to build and operate robust data pipelines in Paris. The role focuses on batch and streaming processing, data lake storage, and producing high-quality, observable datasets.

You will optimize queries, validate data quality, and integrate production experiments within amo’s systems. The ideal candidate has strong Python/SQL skills, production experience with Python or Scala, and hands-on experience with PySpark, Spark SQL, Spark Streaming, Parquet and Iceberg.

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

Senior Data Pipeline Engineer (Batch & Streaming)
Senior Data Pipeline Engineer (Batch & Streaming)

amo • Paris

Sur place
EUR 70 000 - 110 000
Onsite 5 days a week
Central Paris office near Opera
Senior Data Engineer – Spark, Databricks & AI Pipelines (Hybrid)
Senior Data Engineer – Spark, Databricks & AI Pipelines (Hybrid)

Odaseva • Paris

Hybride
EUR 95 000 - 130 000
Paris Data Engineer: Real-Time Pipelines & Trusted Data
Paris Data Engineer: Real-Time Pipelines & Trusted Data

The French Sourcer • Paris

Sur place
EUR 75 000 - 110 000
Health insurance
Meal vouchers
Remote-friendly policy
Senior Data Engineer: Build Scalable Data Pipelines & AI
Senior Data Engineer: Build Scalable Data Pipelines & AI

Technology & Strategy • Paris

Sur place
EUR 45 000 - 60 000
Senior Data Engineer: AI-Powered Data Pipelines
Senior Data Engineer: AI-Powered Data Pipelines

Voodoo • Paris

Hybride
EUR 90 000 - 140 000
Competitive salary based on experience
Swile Lunch voucher
Gymlib (100% covered by Voodoo)
+3
Data Engineer: Scalable Pipelines & AI-Driven Data
Data Engineer: Scalable Pipelines & AI-Driven Data

ManoMano • Paris

Sur place
EUR 55 000 - 75 000
Data Platform Engineer — PySpark, Lakehouse & Databricks
Data Platform Engineer — PySpark, Lakehouse & Databricks

Launchmetrics • Paris

Hybride
EUR 90 000 - 120 000
Learning & development allowance
Benefits package tailored to location
Flexible working arrangements
+1
Data Engineer Team Lead: Pipelines & Analytics
Data Engineer Team Lead: Pipelines & Analytics

SCOR Group • Paris

Hybride
EUR 70 000 - 110 000
Senior Data Engineer: Real-Time Pipelines & Spark
Senior Data Engineer: Real-Time Pipelines & Spark

Hashlist • Poissy

Sur place
EUR 85 000 - 125 000
Senior Data Engineer
Senior Data Engineer

Hashlist • Poissy

Sur place
EUR 85 000 - 125 000