Research Intern

mercor

San Francisco (CA)

On-site

USD 17,000 - 20,000

Full time

3 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Mentorship from experienced researcher
Meal stipend $1.5K monthly for meals
Laundry reimbursement $200 monthly
Wellness reimbursement $200 monthly
Equinox membership
Team events and offsites
Potential full-time return offer

Job summary

Mercor’s Research Scientist Intern role focuses on frontier research in RLVR, data generation, and model evaluation. You will design controlled experiments to explore how datasets and training methods impact large language models, and collaborate with researchers to turn open questions into rigorous studies.

You’ll work with research scientists, engineers, and domain experts to advance Mercor’s agenda, potentially contributing to publications, benchmarks, and frontier AI systems.

Qualifications

  • Demonstrated ability to formulate research questions, design experiments, and draw sound conclusions from empirical results.
  • Experience training, fine‑tuning, or evaluating language models.
  • Familiarity with RL environments, reward modeling, or agentic AI systems.
  • Developing benchmarks, evaluation methodologies, or data-quality measures.
  • At least one publication or open source project.
  • Strong programming skills, particularly in Python, and the ability to write reliable research code.
  • Familiarity with machine learning fundamentals, experimental design, and statistical analysis.
  • Intellectual curiosity and ownership in a fast-paced research environment.

Responsibilities

  • Develop and investigate research questions related to post‑training, RLVR, data quality, and model evaluation.
  • Design and run controlled experiments to understand how datasets, rewards, and training strategies affect model performance.
  • Study reward‑shaping and post‑training methods, including approaches such as GRPO and DAPO.
  • Develop methods for measuring data quality, usability, and performance uplift on key benchmarks.
  • Design and evaluate datasets, rubrics, evaluators, and scoring frameworks for complex model capabilities.
  • Conduct systematic error analysis to identify model failure modes and opportunities for improvement.
  • Analyze experimental results and communicate findings through clear reports, research artifacts, and presentations.
  • Build the research tooling and data pipelines needed to conduct experiments at scale.
  • Collaborate with research scientists, research engineers, applied AI teams, and domain experts producing training and evaluation data.
  • Contribute to research publications, benchmark releases, and other public research outputs where appropriate.

Skills

Research design
Python programming
Empirical analysis
Publications or OSS
RL environments
Language models evaluation
Intellectual curiosity
Fast-paced environment

Education

Master's or PhD in CS/ML/Statistics

Tools

Python

Job description

About Mercor

Mercor's mission is to organize human intelligence to power the AI economy. We're a leading AI data company, building the layer between human expertise and frontier models. Millions of domain experts on the platform are paid over $4 million per day to train frontier AI models. Mercor's APEX benchmark family measures AI's real-world impact on professional work. Mercor Enterprise brings this same infrastructure to Fortune 500 companies: helping companies capture how their best people actually work, translating that expertise directly back into agents.

Mercor is creating a new category of work where expertise powers AI advancement. Achieving this requires an ambitious, fast-paced and deeply committed team. You'll work alongside researchers, operators, and AI companies at the forefront of shaping the systems that are redefining society. Mercor is a profitable Series C company valued at $10 billion. We work in-person five days a week in our San Francisco, NYC, or London offices.

About the Role

As a Research Scientist Intern at Mercor, you'll work on research at the frontier of post-training, reinforcement learning with verifiable rewards (RLVR), data generation, and model evaluation.

You’ll investigate how datasets, rewards, and training methods affect the capabilities and behavior of large language models. This may include designing controlled experiments, developing new evaluation methodologies, conducting systematic failure analysis, and testing approaches to improve tool use, agentic behavior, and real-world reasoning.

You’ll work closely with research scientists, research engineers, and domain experts to turn open-ended questions into rigorous experiments. Your work will contribute to Mercor's research agenda and may support external publications, benchmark releases, and the development of frontier AI systems.

What You'll Do
  • Develop and investigate research questions related to post‑training, RLVR, data quality, and model evaluation.
  • Design and run controlled experiments to understand how datasets, rewards, and training strategies affect model performance.
  • Study reward‑shaping and post‑training methods, including approaches such as GRPO and DAPO.
  • Develop methods for measuring data quality, usability, and performance uplift on key benchmarks.
  • Design and evaluate datasets, rubrics, evaluators, and scoring frameworks for complex model capabilities.
  • Conduct systematic error analysis to identify model failure modes and opportunities for improvement.
  • Analyze experimental results and communicate findings through clear reports, research artifacts, and presentations.
  • Build the research tooling and data pipelines needed to conduct experiments at scale.
  • Collaborate with research scientists, research engineers, applied AI teams, and domain experts producing training and evaluation data.
  • Contribute to research publications, benchmark releases, and other public research outputs where appropriate.
What We're Looking For

Currently pursuing a master's or PhD in computer science, machine learning, statistics, mathematics, or another relevant field.

  • Demonstrated ability to formulate research questions, design experiments, and draw sound conclusions from empirical results.
  • Demonstrated experience in at least one of the following:
  • Training, fine‑tuning, or evaluating language models.
  • Agentic AI system, RL environments.
  • Developing benchmarks, evaluation methodologies, or data‑quality measures.
  • At least one publication or open source project.
  • Strong programming skills, particularly in Python, and the ability to write reliable research code.
  • Familiarity with machine learning fundamentals, experimental design, and statistical analysis.
  • Intellectual curiosity.
  • Comfort operating in a fast‑paced research environment with rapid iteration and a high degree of ownership.
Nice to Have
  • Previous research experience in language models, reinforcement learning, model evaluation, or post‑training.
  • Experience training, fine‑tuning, or evaluating language models.
  • Familiarity with RLVR techniques, reward modeling, or agentic AI systems.
  • Experience developing benchmarks, evaluation methodologies, or data‑quality measures.
  • Research publications or submissions at competitive CS conferences such as ACL, NeurIPS, ICML, ICLR, or EMNLP.
  • Research papers, technical reports, open‑source projects, or other work samples demonstrating relevant skills.
Why Mercor
  • Impact : Your work powers how AI labs train and deploy their models
  • Learning : Get early exposure to frontier AI research and engineering
  • Growth : Work with a high-velocity team where interns ship to production
Benefits
  • Mentorship from experienced researchers.
  • Work on real, high-impact projects.
  • $1.5K monthly stipend for meals
  • $200 monthly laundry reimbursement
  • $200 monthly personal wellness reimbursement
  • Free Equinox membership
  • Team events and offsites.
  • Potential full-time return offer.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Engineer - Environments, Data and Post-Training
Research Engineer - Environments, Data and Post-Training

AI Chopping Block • San Francisco (CA), Northern (KY)

On-site
USD 150,000 - 230,000
Bi-annual performance bonus
Generous equity grant
Relocation bonus up to $15k
+6
Research Engineer - Environments, Data and Post-Training
Research Engineer - Environments, Data and Post-Training

Mercor • San Francisco (CA)

On-site
USD 150,000 - 210,000
Generous equity grant
$10K housing bonus
$1.5K monthly food stipend
+2
Research Engineer – Benchmarking
Research Engineer – Benchmarking

Mercor • San Francisco (CA)

On-site
USD 150,000 - 210,000
Bi-annual bonus
Equity grant
Relocation bonus
+6
Research Operations, Code
Research Operations, Code

Mercor • San Francisco (CA)

On-site
USD 120,000 - 180,000
Bi-annual performance bonus
Equity grant
Relocation bonus
+6
Software Engineer Intern
Software Engineer Intern

Mercor • San Francisco (CA), Northern (KY)

Hybrid
USD 15,000 - 18,000
Mentorship from engineers
Meal stipend of $1,500/month
Equinox membership
+2
Data Science Intern
Data Science Intern

Mercor • San Francisco (CA), Northern (KY)

Hybrid
USD 34,000 - 55,000
Bi-annual performance bonus structure
Generous equity grant vested over 4年
Up to $15k Relocation bonus
+6
Software Engineer — AI Agents & RL Platforms
Software Engineer — AI Agents & RL Platforms

Mercor • San Francisco (CA)

On-site
USD 170,000 - 260,000
Software Engineer, Applied AI - Frontier Engineering (NYC)
Software Engineer, Applied AI - Frontier Engineering (NYC)

Engg • New York (NY)

On-site
USD 180,000 - 240,000
Relocation bonus
Proximity bonus
Meals stipend
+5
Software Engineer, Agents
Software Engineer, Agents

Mercor Inc. • New York (NY), Northern (KY)

On-site
USD 150,000 - 210,000
Bi-annual performance bonus structure
Generous equity grant vested over 4 4 
Up to $15k Relocation bonus
+6
Software Engineer, Search Systems - Code Data
Software Engineer, Search Systems - Code Data

Mercor • San Francisco (CA)

On-site
USD 250,000 - 500,000
Relocation bonus
Housing bonus
Meals stipend
+3