Remote Quantitative Analyst for GenAI Benchmarking

Mercor

New York (NY)

Remote

USD 100,000 - 180,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Mercor platform through Cincinnatus LLC seeks experienced data scientists and quantitative analysts for ground-truth tasks. You will design analytical challenges, author notebooks, and compare methods while evaluating model performance.

This is a full-time W-2 role, remote across the United States, ~35 hours weekly, placed within a leading AI lab. You will work with researchers to ensure rigorous analysis, clear reporting, and practical recommendations for frontier-model evaluations.

Qualifications

  • MSc/PhD in statistics, data science, or quantitative STEM, or equivalent experience.
  • 1+ years in a research, research-engineering, or heavy data-analysis role.
  • Hands-on data-analysis skills: data cleaning, correlation, hypothesis testing, interpretation.
  • Proficiency with notebooks for analysis and reporting; Python and Git.

Responsibilities

  • Design tasks: create realistic data-analysis challenges and cleaning workflows.
  • Author notebooks: produce reproducible reference analyses in notebooks.
  • Compare methods: build tasks with fair comparisons and clear recommendations.
  • Evaluate models: assess how models handle tasks and whether conclusions hold.
  • Work as a team: align with researchers to ensure consistent, accurate evaluations.

Skills

Data analysis
Statistical testing
Python programming
Data cleaning
Communication writing
Independent work
Ambiguity tolerance
Availability 35h/week

Education

MSc or PhD in statistics, data science, or quantitative STEM field
Equivalent practical experience in a research-heavy analytical domain

Tools

Jupyter Notebooks
Google Colab
Git

Job description

Mercor platform through Cincinnatus LLC seeks experienced data scientists and quantitative analysts for ground-truth tasks. You will design analytical challenges, author notebooks, and compare methods while evaluating model performance.

This is a full-time W-2 role, remote across the United States, ~35 hours weekly, placed within a leading AI lab. You will work with researchers to ensure rigorous analysis, clear reporting, and practical recommendations for frontier-model evaluations.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote Quantitative Analyst for AI Benchmarking
Remote Quantitative Analyst for AI Benchmarking

Mercor • New York (NY)

Remote
USD 90,000 - 130,000
Remote GenAI Benchmark Architect — Data Science
Remote GenAI Benchmark Architect — Data Science

Mercor • New York (NY)

On-site
USD 120,000 - 170,000
Remote Data Scientist – GenAI Benchmark & Task Design
Remote Data Scientist – GenAI Benchmark & Task Design

Obsidian • New York (NY)

On-site
USD 90,000 - 150,000
Remote AI Benchmark Test Engineer
Remote AI Benchmark Test Engineer

Mercor • New York (NY)

Remote
USD 85,000 - 120,000
GenAI Benchmark Research Scientist (Remote, 35h/wk)
GenAI Benchmark Research Scientist (Remote, 35h/wk)

Obsidian • San Francisco (CA)

On-site
USD 100,000 - 160,000
GenAI Benchmark Research Scientist - Remote (35h/wk)
GenAI Benchmark Research Scientist - Remote (35h/wk)

Mercor • San Francisco (CA)

Remote
USD 120,000 - 180,000
Senior Software Domain Expert — GenAI QA & Benchmarks
Senior Software Domain Expert — GenAI QA & Benchmarks

Mercor • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
GenAI Benchmark Research Scientist (Remote, Part-Time)
GenAI Benchmark Research Scientist (Remote, Part-Time)

Obsidian • San Francisco (CA)

Remote
USD 120,000 - 160,000
Data Science Expert - AI/ML
Data Science Expert - AI/ML

Obsidian • San Francisco (CA)

On-site
USD 100,000 - 160,000
GenAI Benchmark Research Scientist — Remote, Part-Time
GenAI Benchmark Research Scientist — Remote, Part-Time

Obsidian • New York (NY)

Remote
USD 120,000 - 150,000