Data Engineer - AI Model Evaluation

Mercor

Berlin

Vor Ort

EUR 393.000 - 560.000

Vollzeit

14 Tage+
Bewerbungsgenerator

A complete application in a minute — tailored resume and cover letter, ready to send.

Schaffe es an den ATS-Filtern vorbei

Zusammenfassung

Mercor, in partnership with a leading AI research lab, seeks contributors to evaluate frontier AI coding models through structured technical assessments. You will work on realistic data engineering workflows and model evaluation.

This sprint-based engagement focuses on ETL pipelines, data warehouses, analytics platforms, and distributed data systems, with compensation per accepted task.

Tasks typically take 2–3 hours after ramp-up, and compensation is $400 per accepted task.

Qualifikationen

  • 2+ years of professional data engineering experience.
  • Experience building ETL pipelines, data warehouses, analytics platforms, or distributed data systems.
  • Regular use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI, or similar tools.
  • Ability to evaluate model-generated data infrastructure and pipeline implementations.
  • Experience operating large-scale data platforms is preferred.

Aufgaben

  • Use frontier AI coding agents to complete and evaluate complex data engineering tasks.
  • Review model-generated implementations involving ETL pipelines, data warehouses, analytics platforms, and distributed data systems.
  • Identify bugs, edge cases, scalability issues, and failure modes.
  • Compare outputs from multiple frontier models and assess their strengths and weaknesses.
  • Apply professional engineering judgment to realistic data engineering scenarios.

Kenntnisse

Data engineering
ETL pipelines
Data warehouses
Analytics platforms
Distributed data systems
AI coding agents

Tools

Cursor
Claude Code
Codex
Windsurf
Gemini CLI

Jobbeschreibung

About the Role
  • Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project.
  • Contributors help evaluate and improve frontier AI coding models through structured technical assessments.
  • The work focuses on realistic data engineering workflows and model evaluation.
  • Spots are limited and filling quickly on a first come, first serve basis.
What You'll Do
  • Use frontier AI coding agents to complete and evaluate complex data engineering tasks.
  • Review model-generated implementations involving ETL pipelines, data warehouses, analytics platforms, and distributed data systems.
  • Identify bugs, edge cases, scalability issues, and failure modes.
  • Compare outputs from multiple frontier models and assess their strengths and weaknesses.
  • Apply professional engineering judgment to realistic data engineering scenarios.
Time Commitment
  • Sprint based project that runs in 12-24 hour stretches based on client requirement.
Compensation
  • $400 per accepted task.
  • Typical tasks take approximately 2-3 hours after ramp-up.
  • Compensation is tied to accepted work.
Who Should Apply
  • 2+ years of professional data engineering experience.
  • Experience building ETL pipelines, data warehouses, analytics platforms, or distributed data systems.
  • Regular use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI, or similar tools.
  • Ability to evaluate model-generated data infrastructure and pipeline implementations.
  • Experience operating large-scale data platforms is preferred.
Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

DevOps Engineer - AI Model Evaluator
DevOps Engineer - AI Model Evaluator

Mercor • Berlin

Vor Ort
EUR 15.000 - 31.000
DevOps / SRE / Cloud Engineer (Coding Agent Experience)
DevOps / SRE / Cloud Engineer (Coding Agent Experience)

aitrainer • Deutschland

Vor Ort
EUR 43.066 - 68.906
Security Engineer - Fully Remote | Upto $85/hr
Security Engineer - Fully Remote | Upto $85/hr

Obsidian • Berlin

Remote
EUR 242.493 - 363.740
Risk Engineer - Fully Remote | Upto $80/hr
Risk Engineer - Fully Remote | Upto $80/hr

Obsidian • Berlin

Remote
EUR 242.493 - 484.986
Frontier Engineer (M/F/D)
Frontier Engineer (M/F/D)

Cognizant • Karlsruhe

Hybrid
EUR 90.000 - 130.000
Senior Software Engineer - Agent Evaluation
Senior Software Engineer - Agent Evaluation

aitrainer • Deutschland

Vor Ort
EUR 47.779 - 71.669
Data Science Expert - Evaluation Specialist
Data Science Expert - Evaluation Specialist

Mercor • Berlin

Vor Ort
EUR 90.000 - 120.000
Senior AI Agent Evaluation Engineer
Senior AI Agent Evaluation Engineer

aitrainer • Deutschland

Vor Ort
EUR 34.453 - 60.293
Data Scientist (Python & SQL) - Freelance AI Trainer
Data Scientist (Python & SQL) - Freelance AI Trainer

Mindrift • Stuttgart

Vor Ort
EUR 69.214
Python Engineer, AI Coding Agent Evaluator
Python Engineer, AI Coding Agent Evaluator

g2i • Deutschland

Vor Ort
EUR 119.000 - 238.000