STEM Researchers - Benchmark & First-Author Research

Gramian Consulting Group

United States

Remote

USD 41,000 - 110,000

Part time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Gramian Consultancy is seeking Fellows across STEM fields to help build a new benchmark evaluating frontier AI models for real scientific research. You will reproduce papers, validate AI-generated results, and help define what constitutes good research for AI systems.

Selected Fellows will be named lead/first authors on the benchmark paper and release. This remote, global role offers flexible hours and a contract-based research fellowship arrangement with compensation in the $30–$80/hour range.

Qualifications

  • PhD students or postdocs in STEM fields are welcome to apply.
  • Strong ability to critically review research and identify errors.
  • Experience with AI-assisted scientific research is a plus.

Responsibilities

  • Validate AI-generated paper reproductions, simulations, and research results.
  • Map key research areas and taxonomy within your field.
  • Define what good research looks like for AI systems.
  • Create benchmark tasks and evaluation standards for frontier models.
  • Contribute to the benchmark paper and public evaluation release.
  • Be a lead/first author on a public benchmark paper.
  • Work with a niche AI research lab and frontier AI researchers.

Education

PhD student or postdoc

Tools

Claude Code
Codex

Job description

About Gramian

Gramian Consultancy is a boutique consultancy specializing in IT professional services and engineering talent solutions. With a strong background in software engineering and leadership, we help companies build high-performing teams by matching them with professionals who truly fit their needs.

About the Role

We’re working with a highly specialized AI research lab building new benchmarks for how frontier AI systems perform real scientific research. They are looking for PhD students and postdocs across STEM fields to help reproduce papers, validate AI-generated research, and define what "good research" should look like for AI systems.

We are looking for one Fellow per STEM field to help create a new benchmark for evaluating how well frontier AI models perform real scientific research.

The standout part of this opportunity: the benchmark will be released publicly with a research paper and open evaluation set, and Fellows will be named lead / first authors.

CONTRACT: Research fellowship / contractor

LOCATIONS: Remote, global

COMMITMENT: Flexible hours

COMPENSATION: $30-$80/hour

PROCESS: Application - research review - interview

Fields: Biology, Medicine, Neuroscience, Materials Science, Chemistry, Physics, Mechanical Engineering, Chemical Engineering.

Responsibilities
  • Validate AI-generated paper reproductions, simulations, and research results.
  • Map key research areas and taxonomy within your field.
  • Define what "good research" looks like for AI systems.
  • Create benchmark tasks and evaluation standards for frontier models.
  • Contribute directly to the benchmark paper and public eval release.
  • Current PhD student or postdoc in a relevant STEM field.
  • Based at a strong research university or institute.
  • At least one published paper in your field.
  • Able to critically review research and identify methodological or technical errors.
  • Currently using AI for Science in your own research.
  • Significant hands-on use of tools such as Claude Code, Codex, or similar AI agents.
  • Be a lead / first author on a public benchmark paper.
  • Help create one of the first research benchmarks in your scientific field.
  • Work directly with a niche AI research lab and frontier AI researchers.
  • Flexible remote work designed around academic commitments.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

STEM Researchers - Benchmark & First-Author Research
STEM Researchers - Benchmark & First-Author Research

Gramian Consulting • Massachusetts

Remote
USD 41,328,000 - 110,208,000
Lead author on benchmark paper
Flexible remote work
AI Research Benchmark Fellow: Lead Author (Remote)
AI Research Benchmark Fellow: Lead Author (Remote)

Gramian Consulting Group • United States

Remote
USD 41,000 - 110,000
STEM Researcher - Computational Fields
STEM Researcher - Computational Fields

Weekday 1 • United States

Remote
USD 83,000 - 124,000
Frontier AI Research Fellow — Lead Author (Remote)
Frontier AI Research Fellow — Lead Author (Remote)

Gramian Consulting • Massachusetts

Remote
USD 41,328,000 - 110,208,000
Lead author on benchmark paper
Flexible remote work
Life Sciences Researcher (AI Evaluation)
Life Sciences Researcher (AI Evaluation)

Gramian Consulting • United States

Remote
PKR 1,378,000 - 2,755,000
Applied Computer Science Benchmark Specialist
Applied Computer Science Benchmark Specialist

Weekday 1 • United States

Remote
USD 91,000 - 116,000
Fully remote
Software Engineering Expert
Software Engineering Expert

Weekday 1 • United States

Remote
USD 83,000 - 124,000
Fully remote
Weekly payments
Materials Science Analyst
Materials Science Analyst

Gramian Consulting Group • United States

Remote
USD 83,000 - 165,000
Senior Research Scientist, STEM
Senior Research Scientist, STEM

Foundation Capital • United States

On-site
USD 250,000 - 350,000
Software Engineer, Benchmarking
Software Engineer, Benchmarking

Epoch AI • United States

On-site
USD 125,000 - 200,000
Comprehensive health insurance
Flexible work environment
Generous paid time off
+1