STEM Researchers - Benchmark & First-Author Research

Gramian Consulting

Massachusetts

Remote

USD 41,328,000 - 110,208,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Lead author on benchmark paper
Flexible remote work

Job summary

Gramian Consultancy collaborates with a specialized AI research lab to recruit Fellows across STEM fields to reproduce papers, validate AI-derived research, and define what constitutes good research for frontier AI systems.

Fellows will be named lead/first authors on a public benchmark paper, with flexible remote work designed around academic commitments and a global pool of collaborators.

Qualifications

  • Current PhD student or postdoc in a relevant STEM field.
  • Ability to critically review research and identify methodological or technical issues.
  • Experience using AI for scientific research.

Responsibilities

  • Validate AI-generated paper reproductions, simulations, and research results.
  • Map key research areas and taxonomy within your field.
  • Define what \"good research\" looks like for AI systems.
  • Create benchmark tasks and evaluation standards for frontier models.
  • Contribute to the benchmark paper and public evaluation release.

Skills

Critical review
AI for science

Education

PhD student or postdoc

Tools

Claude Code
Codex

Job description

About Gramian

Gramian Consultancy is a boutique consultancy specializing in IT professional services and engineering talent solutions. With a strong background in software engineering and leadership, we help companies build high-performing teams by matching them with professionals who truly fit their needs.

About Gramian

Gramian Consultancy is a boutique consultancy specializing in IT professional services and engineering talent solutions. With a strong background in software engineering and leadership, we help companies build high-performing teams by matching them with professionals who truly fit their needs.

About The Role

We're working with a highly specialized AI research lab building new benchmarks for how frontier AI systems perform real scientific research. They are looking for PhD students and postdocs across STEM fields to help reproduce papers, validate AI-generated research, and define what "good research" should look like for AI systems. We are looking for one Fellow per STEM field to help create a new benchmark for evaluating how well frontier AI models perform real scientific research. The standout part of this opportunity: the benchmark will be released publicly with a research paper and open evaluation set, and Fellows will be named lead / first authors.

CONTRACT

Research fellowship / contractor

LOCATIONS

Remote, global

COMMITMENT

Flexible hours

COMPENSATION

$30-$80/hour

PROCESS

Application → research review → interview

Fields

Biology, Medicine, Neuroscience, Materials Science, Chemistry, Physics, Mechanical Engineering, Chemical Engineering.

Responsibilities
  • Validate AI-generated paper reproductions, simulations, and research results.
  • Map key research areas and taxonomy within your field.
  • Define what "good research" looks like for AI systems.
  • Create benchmark tasks and evaluation standards for frontier models.
  • Contribute directly to the benchmark paper and public eval release
Requirements
  • Current PhD student or postdoc in a relevant STEM field.
  • Based at a strong research university or institute.
  • At least one published paper in your field.
  • Able to critically review research and identify methodological or technical errors.
  • Currently using AI for Science in your own research.
  • Significant hands-on use of tools such as Claude Code, Codex, or similar AI agents
Benefits
  • Be a lead / first author on a public benchmark paper.
  • Help create one of the first research benchmarks in your scientific field.
  • Work directly with a niche AI research lab and frontier AI researchers.
  • Flexible remote work designed around academic commitments
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

STEM Researchers - Benchmark & First-Author Research
STEM Researchers - Benchmark & First-Author Research

Gramian Consulting Group • United States

Remote
USD 41,000 - 110,000
AI Research Benchmark Fellow: Lead Author (Remote)
AI Research Benchmark Fellow: Lead Author (Remote)

Gramian Consulting Group • United States

Remote
USD 41,000 - 110,000
STEM Researcher - Computational Fields
STEM Researcher - Computational Fields

Weekday 1 • United States

Remote
USD 83,000 - 124,000
Frontier AI Research Fellow — Lead Author (Remote)
Frontier AI Research Fellow — Lead Author (Remote)

Gramian Consulting • Massachusetts

Remote
USD 41,328,000 - 110,208,000
Lead author on benchmark paper
Flexible remote work
Life Sciences Researcher (AI Evaluation)
Life Sciences Researcher (AI Evaluation)

Gramian Consulting • United States

Remote
PKR 1,378,000 - 2,755,000
Software Engineering Expert
Software Engineering Expert

Weekday 1 • United States

Remote
USD 83,000 - 124,000
Fully remote
Weekly payments
Applied Computer Science Benchmark Specialist
Applied Computer Science Benchmark Specialist

Weekday 1 • United States

Remote
USD 91,000 - 116,000
Fully remote
AI Benchmark & Datasets Engineer / Researcher
AI Benchmark & Datasets Engineer / Researcher

Ignite Next GmbH • Palo Alto (CA), Northern (KY)

Hybrid
USD 120,000 - 190,000
Remote work
Office visits Palo Alto, Paris, Wroclw
Staff Research Scientist
Staff Research Scientist

Foundation Capital • United States

On-site
USD 250,000 - 400,000
Senior Research Scientist, STEM
Senior Research Scientist, STEM

Foundation Capital • United States

On-site
USD 250,000 - 350,000