Remote Quantitative Analyst for GenAI Benchmarking

Mercor

New York (NY)

Remote

USD 100,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Mercor platform through Cincinnatus LLC seeks experienced data scientists and quantitative analysts for ground-truth tasks. You will design analytical challenges, author notebooks, and compare methods while evaluating model performance.

This is a full-time W-2 role, remote across the United States, ~35 hours weekly, placed within a leading AI lab. You will work with researchers to ensure rigorous analysis, clear reporting, and practical recommendations for frontier-model evaluations.

Qualifications

  • MSc/PhD in statistics, data science, or quantitative STEM, or equivalent experience.
  • 1+ years in a research, research-engineering, or heavy data-analysis role.
  • Hands-on data-analysis skills: data cleaning, correlation, hypothesis testing, interpretation.
  • Proficiency with notebooks for analysis and reporting; Python and Git.

Responsibilities

  • Design tasks: create realistic data-analysis challenges and cleaning workflows.
  • Author notebooks: produce reproducible reference analyses in notebooks.
  • Compare methods: build tasks with fair comparisons and clear recommendations.
  • Evaluate models: assess how models handle tasks and whether conclusions hold.
  • Work as a team: align with researchers to ensure consistent, accurate evaluations.

Skills

Data analysis
Statistical testing
Python programming
Data cleaning
Communication writing
Independent work
Ambiguity tolerance
Availability 35h/week

Education

MSc or PhD in statistics, data science, or quantitative STEM field
Equivalent practical experience in a research-heavy analytical domain

Tools

Jupyter Notebooks
Google Colab
Git

Job description

Mercor platform through Cincinnatus LLC seeks experienced data scientists and quantitative analysts for ground-truth tasks. You will design analytical challenges, author notebooks, and compare methods while evaluating model performance.

This is a full-time W-2 role, remote across the United States, ~35 hours weekly, placed within a leading AI lab. You will work with researchers to ensure rigorous analysis, clear reporting, and practical recommendations for frontier-model evaluations.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote Quantitative Analyst for AI Benchmarking
Remote Quantitative Analyst for AI Benchmarking

Mercor • New York (NY)

Remote
USD 90,000 - 130,000
GenAI Benchmark Research Scientist (Remote, 35h/wk)
GenAI Benchmark Research Scientist (Remote, 35h/wk)

Obsidian • San Francisco (CA)

On-site
USD 100,000 - 160,000
Data Science & Quantitative Analysis Expert
Data Science & Quantitative Analysis Expert

Dorado • United States

Remote
USD 90,000 - 120,000
Remote Data Science & Quant Analytics Expert
Remote Data Science & Quant Analytics Expert

Dorado • United States

Remote
USD 90,000 - 120,000
GenAI Benchmark Research Scientist - Remote (35h/wk)
GenAI Benchmark Research Scientist - Remote (35h/wk)

Mercor • San Francisco (CA)

Remote
USD 120,000 - 180,000
GenAI Benchmark Research Scientist (Remote, Part-Time)
GenAI Benchmark Research Scientist (Remote, Part-Time)

Obsidian • San Francisco (CA)

Remote
USD 120,000 - 160,000
GenAI Benchmark Designer - Remote Research (35h/wk)
GenAI Benchmark Designer - Remote Research (35h/wk)

Dorado • United States

Remote
USD 105,000 - 150,000
Data Science Expert - AI/ML
Data Science Expert - AI/ML

Obsidian • San Francisco (CA)

On-site
USD 100,000 - 160,000
Remote QA/Test Engineer for AI Benchmarks
Remote QA/Test Engineer for AI Benchmarks

Dorado • United States

Remote
USD 90,000 - 130,000
GenAI Model Evaluation Engineer — Remote, 35h/wk
GenAI Model Evaluation Engineer — Remote, 35h/wk

Dorado • United States

Remote
USD 120,000 - 180,000