AI Benchmark Scientist - Quantum & Computational Chemistry

Obsidian

Greater London

Remote

GBP 41,000 - 69,000

Part time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Mercor is seeking PhD- and Master’s-level scientists to author AI evaluation tasks as part of a new Sci Code benchmark. You will craft original, executable research problems that current frontier models cannot solve and contribute to a benchmark for scientific computing with a focus on chemistry subdomains.

You will source material, write prompts, define grading criteria, and calibrate tasks against models that fail more often than succeed, in a collaborative experiment with leading AI labs.

Qualifications

  • PhD in chemistry, physical chemistry, theoretical chemistry, computational chemistry, or a closely related field.
  • Demonstrated depth in quantum chemistry and computational chemistry.
  • Proficiency in Python, R, or another language for scientific computing.
  • Comfortable with Git/GitHub and running code in Docker — PR workflow experience.

Responsibilities

  • Source your own material from papers, Kaggle datasets, open-source repositories, or a scenario you design.
  • Write scientific prompts based on the input.
  • Build the grading criteria that define a correct answer.
  • Calibrate against frontier models — a task ships only when strong models fail it more often than they succeed.

Skills

Python
R
Git/GitHub
Docker

Education

PhD in chemistry or closely related field

Tools

Docker
Git/GitHub

Job description

Mercor is seeking PhD- and Master’s-level scientists to author AI evaluation tasks as part of a new Sci Code benchmark. You will craft original, executable research problems that current frontier models cannot solve and contribute to a benchmark for scientific computing with a focus on chemistry subdomains.

You will source material, write prompts, define grading criteria, and calibrate tasks against models that fail more often than succeed, in a collaborative experiment with leading AI labs.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Quantum Chemistry AI Benchmark Architect (Part-Time, 6 Wks)
Quantum Chemistry AI Benchmark Architect (Part-Time, 6 Wks)

Mercor • Greater London

On-site
GBP 41,000 - 69,000
AI Benchmark Scientist: Biochemistry & Genetics
AI Benchmark Scientist: Biochemistry & Genetics

Obsidian • Greater London

Remote
GBP 40,000 - 50,000
Quantum Chemistry AI Research Engineer (Part-Time)
Quantum Chemistry AI Research Engineer (Part-Time)

Obsidian • Greater London

On-site
GBP 55,000 - 83,000
Computational Chemist: AI Benchmark Developer (6-Week)
Computational Chemist: AI Benchmark Developer (6-Week)

Mercor • Greater London

Remote
GBP 6,612,000 - 9,919,000
Quantum Computing AI Benchmark Designer
Quantum Computing AI Benchmark Designer

Mercor • Greater London

On-site
GBP 273,000 - 455,000
AI Benchmark Scientist: Biology PhD Coder
AI Benchmark Scientist: Biology PhD Coder

Mercor • Greater London

On-site
GBP 55,000 - 83,000
Quantum Computing Scientist AI Benchmarking & Prompt Design
Quantum Computing Scientist AI Benchmarking & Prompt Design

Obsidian • Greater London

Remote
GBP 7,200 - 11,000
Physics PhD — AI Benchmark Scientist
Physics PhD — AI Benchmark Scientist

Mercor • Greater London

On-site
GBP 39,000 - 67,000
Computational Mathematician: AI Benchmark Problem Designer
Computational Mathematician: AI Benchmark Problem Designer

Mercor • Greater London

Remote
GBP 109,000 - 218,000
Physics AI Benchmark Scientist: Prompt Design & Evaluation
Physics AI Benchmark Scientist: Prompt Design & Evaluation

Mercor • Greater London

Remote
GBP 11,000 - 18,000