Senior Data Engineer

Bespoke Labs

Netherlands

Remote

EUR 119,000 - 179,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Hyper-competitive hourly rate
Opportunity to work on cutting-edge AI projects

Job summary

A leading AI research lab is seeking a Senior/Staff Data Engineer for a high-impact contract role. You will architect and develop systems for AI model training using your extensive data engineering experience with platforms such as Python and Spark. Ideal candidates will have a proven track record of managing production data setups within top-tier enterprises and possess a deep understanding of data architecture and processing at scale. This is a remote contract position requiring full-time commitment.

Qualifications

  • 6+ years of experience in Data Engineering.
  • Demonstrated Senior/Staff-level ownership of data platforms.
  • Background at Tier-1 enterprises is essential.

Responsibilities

  • Design the data architecture for large-scale AI model training.
  • Develop custom ingestion logic and transformation scripts.
  • Optimize processing workloads to meet throughput targets.
  • Ensure data quality and reliability throughout the pipeline.
  • Act as technical authority on distributed systems and cloud structures.

Skills

Data Engineering expertise
Python/Scala proficiency
Kafka knowledge
Data architecture design
High-throughput processing
Reliability engineering

Tools

Apache Spark
Airflow
Snowflake
BigQuery
Redshift
Apache Iceberg
Delta Lake

Job description

Location: Remote

Role Type: Contract (2-4 Months)

Time Commitment: 40 hrs/week (Full-time availability required)

Compensation: Hyper-competitive hourly rate (matching Tier-1 Staff engineering bands) Experience: 6-12+ years

About BespokeLabs

BespokeLabs is a premier, VC-backed AI Research lab with an exceptionally talent-dense team of IIT and Ivy League alumni. We don’t just build tooling around AI—we build the massive-scale data systems and reasoning architectures that directly power next-generation models. Our research shapes the frontier of AI: we’ve published breakthroughs like GEPA, driven foundational datasets like OpenThoughts, and shipped state-of-the-art models including Bespoke-MiniCheck and Bespoke-MiniChart. More on our website https://www.bespokelabs.ai/ :)

Role Overview

We are looking for a top-tier Senior/Staff Data Engineer for a high-impact, 2-month sprint. You will leverage your deep expertise in enterprise-grade data platforms to architect and build the complex curation systems required for advanced AI model training.

This is not a traditional ETL pipeline role. We need a heavy-hitter who has already operated production data platforms at scale inside large, complex organizations (FAANG, Fortune 100). You will use the mental models, architectural intuition, and coding skills you've developed over your career to generate, transform, and evaluate the data that trains the next generation of AI.

What You Will Do (The Contract)
  • Architect AI-Scale Systems: Design the overarching data architecture and processing topology needed to programmatically curate and shape datasets at TB/PB scale, ensuring low latency and high consistency.
  • Hands-On Development: Write production-grade code (Python/Scala, Spark) to build custom ingestion logic, highly efficient transformation scripts, and performant data validation checks.
  • Complex Data Logic: Implement advanced filtering, deduplication, and quality-scoring algorithms at scale, ensuring the resulting data objects are optimized for LLM/ML consumption.
  • Quality & Performance Tuning: Rigorously test, benchmark, and optimize processing workloads (CPU/memory tuning, partitioning strategies in Spark/Iceberg) to meet aggressive throughput targets.
  • Domain Subject Matter Expert: Act as the ultimate technical authority on distributed systems, data processing, and cloud structures to ensure the training data factory meets enterprise-grade accuracy.
What You Bring to the Table (Your Past Experience)

To be successful in this contract, you must have a track record of:

  • End-to-End Ownership: Designing and owning enterprise data platforms (batch + streaming).
  • High-Throughput Processing: Building and operating Kafka-first streaming pipelines.
  • Lakehouse Architecture: Utilizing Apache Iceberg, Delta Lake, or Hudi for analytics and ML at scale.
  • Reliability Engineering: Ensuring data reliability through SLAs, monitoring, backfills, and recovery.
  • Scale: Processing billions of events and managing TB–PB scale data systems.
Required Qualifications (Non-Negotiable)
  • Experience: 6+ years of Data Engineering experience.
  • Seniority: Demonstrated Senior/Staff-level ownership of production data platforms.
  • Pedigree: Background at Tier-1 enterprises (FAANG, large SaaS, Fortune 100).
  • Technical Stack: Deep fluency in Python/Scala, Spark, Kafka, Airflow, and Major Cloud Warehouses (Snowflake, BigQuery, Redshift).
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Scientist
Data Scientist

Bespoke Labs • Netherlands

Remote
EUR 165,000 - 248,000
Hyper-competitive hourly rate
Medior/Senior Data Engineer
Medior/Senior Data Engineer

XITE Netherlands • Amsterdam

On-site
EUR 60,000 - 80,000
Senior Data Engineer
Senior Data Engineer

Michael Bailey Associates • Randstad

Hybrid
EUR 73,000 - 87,000
Lease car
Generous learning budget
DevOps Engineer
DevOps Engineer

Bespoke Labs • Netherlands

Remote
Fully remote work
Flexible schedule
Senior Data Engineer ID82269
Senior Data Engineer ID82269

AgileEngine, LLC. • Netherlands

Remote
BRL 180,000 - 240,000
Growth without limits
Competitive compensation
Flexibility: 100% remote
+3
Senior Data Engineer
Senior Data Engineer

Levy Professionals • Amsterdam

On-site
EUR 90,000 - 130,000
Data Engineer
Data Engineer

FRG Technology Consulting • Haarlem

Hybrid
EUR 60,000 - 90,000
Competitive compensation aligned with experience
Collaborative environment
High level of ownership in projects
AI-Ready Data Engineer: Build Scalable Data Pipelines
AI-Ready Data Engineer: Build Scalable Data Pipelines

McKinsey & Company • Amsterdam

On-site
Staff Data Platform Engineer
Staff Data Platform Engineer

Harnham • Amsterdam

On-site
EUR 90,000 - 120,000
Competitive salary
Modern data & ML technologies
Autonomy & ownership
Senior Data Engineer — AI-Scale Systems (Remote, Contract)
Senior Data Engineer — AI-Scale Systems (Remote, Contract)

Bespoke Labs • Netherlands

Remote
EUR 119,000 - 179,000
Hyper-competitive hourly rate
Opportunity to work on cutting-edge AI projects