Data Engineer

The French Sourcer

Paris

Sur place

EUR 75 000 - 110 000

Plein temps

14 jours+

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Avantages offerts par ce poste

Health insurance
Meal vouchers
Remote-friendly policy

Résumé du poste

The French Sourcer is seeking a Data Engineer to join our Paris data platform. You will own batch and streaming pipelines, ingest data from multiple sources, and design robust schemas for growing product lines across markets.

Ideal candidates have 4+ years of experience and strong SQL/Python skills, plus hands-on with Airflow, dbt, Kafka, Snowflake or BigQuery. You will collaborate with analysts and data scientists to deliver trusted datasets and scale performance.

Qualifications

  • 4+ years of experience building and operating production data pipelines.
  • Strong SQL and Python skills.
  • Practical experience with Airflow, dbt, Kafka, Snowflake, or similar technologies.
  • Good knowledge of data modelling and incremental processing.
  • Experience with automated testing, monitoring, and production troubleshooting.
  • Understanding of retries, partial failures, late-arriving data, backfills, and pipeline dependencies.
  • Experience using Git, code reviews, and CI/CD.
  • Confidence working with analysts, data scientists, and engineering teams.
  • Ownership of data quality, not just data movement between systems.

Responsabilités

  • Building and owning batch and streaming pipelines that feed the company’s core data warehouse.
  • Ingesting data from transactional databases, internal services, APIs, and event streams.
  • Designing schemas and data models that remain consistent as the business adds new products and markets.
  • Developing and maintaining reusable transformation models in dbt.
  • Setting up automated checks for freshness, completeness, duplicates, schema changes, and business consistency.
  • Monitoring production pipelines and investigating failures across sources, orchestration, transformations, and warehouse workloads.
  • Managing retries, late-arriving data, backfills, and historical reprocessing.
  • Optimizing pipeline performance, query execution, storage, and compute costs as data volumes increase.
  • Planning schema migrations without disrupting downstream users.
  • Partnering directly with data scientists and analysts to translate their requirements into reliable datasets.
  • Documenting lineage, dependencies, transformation logic, and ownership so critical data can be understood and trusted.
  • Contributing to code reviews, automated testing, CI/CD, and data engineering standards.

Connaissances

SQL
Python
Data pipelines
Data modeling
Troubleshooting
Ownership

Outils

Airflow
dbt
Kafka
Snowflake
BigQuery
Git
CI/CD

Description du poste

Keep millions of daily transactions clean, trustworthy, and moving, without turning every upstream change into a fire drill.

Data Engineer - Paris

Paris, France · Permanent · On-site

What you'd actually work on
  • Building and owning batch and streaming pipelines that feed the company’s core data warehouse
  • Ingesting data from transactional databases, internal services, APIs, and event streams
  • Designing schemas and data models that remain consistent as the business adds new products and markets
  • Developing and maintaining reusable transformation models in dbt
  • Setting up automated checks for freshness, completeness, duplicates, schema changes, and business consistency
  • Monitoring production pipelines and investigating failures across sources, orchestration, transformations, and warehouse workloads
  • Managing retries, late-arriving data, backfills, and historical reprocessing
  • Optimizing pipeline performance, query execution, storage, and compute costs as data volumes increase
  • Planning schema migrations without disrupting downstream users
  • Partnering directly with data scientists and analysts to translate their requirements into reliable datasets
  • Documenting lineage, dependencies, transformation logic, and ownership so critical data can be understood and trusted
  • Contributing to code reviews, automated testing, CI/CD, and data engineering standards
Where it gets technically interesting
  • Streaming ingestion with Kafka alongside batch orchestration with Airflow, with clear decisions about when each approach is appropriate
  • A Snowflake or BigQuery warehouse under real query and transformation load from multiple internal teams
  • A growing dbt transformation layer that needs to remain tested, documented, and maintainable
  • Incremental processing across large datasets where full refreshes are not a realistic option
  • Backfills and schema migrations on tables that cannot simply be rerun from scratch
  • Handling partial failures, upstream changes, delayed events, and dependencies between critical pipelines
  • Improving data quality without producing large volumes of low-value alerts
  • Balancing reliability and performance with infrastructure and warehouse costs
What we're looking for
  • 4+ years of experience building and operating production data pipelines
  • Strong SQL and Python skills
  • Practical experience with Airflow, dbt, Kafka, Snowflake, BigQuery, or similar technologies
  • Good knowledge of data modelling and incremental processing
  • Experience with automated testing, monitoring, and production troubleshooting
  • An understanding of retries, partial failures, late-arriving data, backfills, and pipeline dependencies
  • The ability to investigate issues across the full data flow
  • Experience using Git, code reviews, and CI/CD
  • Confidence working directly with analysts, data scientists, and engineering teams
  • Ownership of data quality, not just the movement of data between systems

You do not need to have worked with every tool in the stack, but you should understand the engineering principles behind reliable and maintainable data systems.

The company

A fast-growing e-commerce platform operating across several European markets, with several hundred employees and millions of transactions processed every month.

The data platform supports reporting, product analytics, finance, operations, and data science. The team is now scaling its pipelines and models to support higher volumes, new markets, and a growing number of internal users.

  • Health insurance, meal vouchers, and a remote-friendly policy.

Languages: Native or bilingual French and professional English.

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

Senior Data Engineer
Senior Data Engineer

Dataworks • Paris

Hybride
EUR 55 000 - 65 000
Hybrid work model
Remote 3 days/week
Autonomy over architecture
Data Engineer
Data Engineer

Enfint • Paris

Sur place
EUR 65 000 - 90 000
Relocation support
Paid time off
Best medical insurance in France
+2
Paris Data Engineer: Real-Time Pipelines & Trusted Data
Paris Data Engineer: Real-Time Pipelines & Trusted Data

The French Sourcer • Paris

Sur place
EUR 75 000 - 110 000
Health insurance
Meal vouchers
Remote-friendly policy
Data Engineer Senior
Data Engineer Senior

GazelTech • Paris

Sur place
EUR 70 000 - 110 000
Data Engineer
Data Engineer

European Tech Recruit • Paris

Hybride
EUR 60 000 - 80 000
Competitive salary + equity
20 days of paid vacation
Hybrid work from Paris
+4
Data Engineer
Data Engineer

5V Tech • Grenoble

Sur place
EUR 65 000 - 90 000
Data Engineer
Data Engineer

Allen Recruitment • Brest

Hybride
EUR 45 000 - 65 000
Data Engineer - AI Safety
Data Engineer - AI Safety

European Tech Recruit • Paris

Sur place
EUR 60 000 - 80 000
Data Engineer
Data Engineer

Teradata • Antony

Sur place
EUR 70 000 - 110 000
Data Engineer H/F - Paris
Data Engineer H/F - Paris

AVISIA • Paris

Sur place
EUR 45 000 - 60 000