STEM Researcher — Computational Fields

Dorado

United States

Remote

USD 105,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Cincinnatus LLC is recruiting for an embedded researcher to join a leading AI lab's GenAI team. The role is fully remote within the United States and ~35 hours per week, with W-2 employment through Cincinnatus LLC and placement within client teams.

You will design studies, implement analyses in Python, and evaluate model outputs, collaborating with researchers to ensure rigorous, reproducible results.

Qualifications

  • MSc or PhD in a STEM field or computational discipline with data analysis and coding.
  • 1+ year in an active research role (academia, industry, or national labs).
  • Research involves substantial computational work: Python-based analysis, modeling, or data pipelines.
  • Strong design, hypothesis testing, and rigorous evaluation skills.
  • Familiar with Git, IDEs, and notebook environments (Jupyter/Colab).
  • Past experience in AI training, model evaluation, or benchmark task authoring preferred.
  • Attention to detail, clear written communication, and ability to work independently.

Responsibilities

  • Design tasks: translate research skills into engaging, multi-step tasks.
  • Author solutions: work through tasks in Python and notebooks with rigor.
  • Define what good looks like: articulate criteria separating strong vs. sound reasoning.
  • Evaluate models: review attempts and flag mistakes a working researcher would spot.
  • Work as a team: compare notes with researchers to ensure consistency.

Skills

Experimental design
Hypothesis testing
Rigorous evaluation
Scientific writing
Independent work
Time management

Education

MSc or PhD in STEM or computational social science / humanities
1+ years in active research role
Computational work: Python, data analysis, coding

Tools

Python
Git
Jupyter/Colab
IDEs

Job description

Join a leading AI lab's cutting-edge GenAI team to be at the core of the AI revolution, where your expertise fuels the development of the most advanced AI models.

1. Overview

A leading AI lab is building the next generation of agentic evaluation benchmarks for frontier models and is recruiting researchers from computational STEM fields — as well as computationally heavy social sciences and humanities — to bring working-researcher rigor to benchmark design. You will translate the scientific method — experimental design, hypothesis testing, and rigorous evaluation — into complex, multi-step tasks that today\'s best models cannot yet complete reliably.

Each task represents one to two days of continuous, focused effort and spans multiple skills: study design, implementation in code, data analysis, and careful written conclusions. You will work in a tight feedback loop with the lab\'s researchers, surfacing the kinds of methodological mistakes a working researcher would catch immediately.

This is a full-time W-2 employment position with Cincinnatus LLC, with the opportunity to be placed at a leading AI lab as part of their extended workforce. This role is fully remote within the United States, at approximately 35 hours per week.

2. Key Responsibilities
  • Design tasks: Turn the research skills you use every day — designing studies, testing hypotheses, evaluating results — into engaging, multi-step tasks.

  • Author solutions: Work through your own tasks in Python and notebooks, at the level of rigor you\'d expect from a careful colleague.

  • Define what good looks like: Help spell out what separates sound scientific reasoning from reasoning that merely sounds right.

  • Evaluate models: Review model attempts at your tasks and flag the mistakes a working researcher would spot right away.

  • Work as a team: Compare notes with researchers and fellow experts to keep evaluations consistent and accurate.

3. Core Qualifications
  • MSc or PhD in a STEM field, or in a computational social-science or humanities discipline, or equivalent practical experience in a research-heavy domain requiring data analysis and coding.

  • 1+ years of experience in an active research role (academia, industry, or national labs).

  • Your own research involves significant computational work: Python-based analysis, simulation, modeling, or data pipelines.

  • Strong grounding in experimental design, hypothesis testing, and rigorous evaluation of results.

  • Working familiarity with Git, IDEs, and notebook environments (Jupyter or Colab).

  • Past experience in AI training, model evaluation, or benchmark/task authoring is preferred.

  • A perfectionist mindset: high attention to detail, creativity in task design, strong written communication, and the ability to work independently through ambiguous, open-ended problems.

  • Ability to engage reliably for approximately 35 hours per week.

About Cincinnatus LLC

Cincinnatus LLC is an enterprise staffing company that partners with leading technology companies to source and employ highly skilled professionals for contingent and contract-based opportunities. Cincinnatus serves as the employer of record for these engagements, providing W-2 employment, payroll, benefits, and compliance, while placing employees directly within client teams to work on high-impact initiatives.

Roles hired through Cincinnatus are not project-based or freelance engagements. They are structured, role-based positions that typically involve part-time or full-time commitments, close collaboration with a client\'s internal teams, and integration into standard enterprise workflows.

Cincinnatus is a legal entity separate from Mercor. While opportunities may be discovered through Mercor\'s platform, employment, onboarding, payroll, and benefits for these roles are administered by Cincinnatus LLC.

Equal Employment Opportunity

Cincinnatus is proud to be an Equal Employment Opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, reproductive health decisions, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, genetic information, political views or activity, or any other legally protected characteristic.

Cincinnatus is committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans throughout the job application process.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Science Expert - AI/ML
Data Science Expert - AI/ML

Obsidian • San Francisco (CA)

On-site
USD 100,000 - 160,000
Data Science & Quantitative Analysis Expert
Data Science & Quantitative Analysis Expert

Dorado • United States

Remote
USD 90,000 - 120,000
Machine Learning Engineer — Model Evaluation & Experimentation
Machine Learning Engineer — Model Evaluation & Experimentation

Dorado • United States

Remote
USD 120,000 - 180,000
Software Domain Expert - Fully Remote
Software Domain Expert - Fully Remote

Mercor • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Software Domain Expert
Software Domain Expert

Mercor • San Francisco (CA)

Hybrid
USD 180,000 - 250,000
Biophysics Researcher - Scientific Expert
Biophysics Researcher - Scientific Expert

Mercor • San Francisco (CA)

On-site
USD 96,000 - 152,000
QA/Test Engineer
QA/Test Engineer

Dorado • United States

Remote
USD 90,000 - 130,000
Engineering & Software Domain Expert Mercor · Bay Area, CA $65-105/hr →
Engineering & Software Domain Expert Mercor · Bay Area, CA $65-105/hr →

Dorado • California (MO), Northern (KY)

Hybrid
USD 180,000 - 240,000
LLM Red Team Specialist — Failure Modes & Edge Cases
LLM Red Team Specialist — Failure Modes & Edge Cases

Dorado • United States

Remote
USD 120,000 - 180,000
Mathematics Researcher - AI Systems
Mathematics Researcher - AI Systems

Obsidian • New York (NY)

On-site
USD 50,000 - 95,000