Materials Scientist: AI Benchmarking & Research Prompts

Obsidian

San Francisco (CA)

Remote

USD 8,266,000 - 12,398,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Mercor is hiring PhD and Masters scientists to author AI evaluation tasks and build a robust benchmark for scientific computing. You will craft original, executable research problems that current frontier models struggle to solve, with focus on semiconductor materials and molecular modeling.

You will source your own material, write prompts, and develop grading criteria while calibrating against failing models.

Qualifications

  • PhD in materials science, materials engineering, applied physics, chemistry, chemical engineering, or a closely related field
  • Demonstrated depth in both semiconductor materials and molecular modeling
  • Working proficiency in Python, R, or another relevant programming language for scientific computing
  • Comfortable with Git/GitHub and running code in Docker — authoring runs through a pull-request workflow with automated quality checks

Responsibilities

  • Author original, executable research problems that today's frontier models cannot solve
  • Build the grading criteria that define a correct answer
  • Calibrate against frontier models — a task ships only when strong models fail it more often than they succeed
  • Write scientific prompts based on the input
  • Source material from papers, datasets, or open-source repositories

Skills

Python
R
Scientific computing
Docker
Git/GitHub
Research design

Education

PhD in materials science or related field

Tools

Docker
GitHub

Job description

Mercor is hiring PhD and Masters scientists to author AI evaluation tasks and build a robust benchmark for scientific computing. You will craft original, executable research problems that current frontier models struggle to solve, with focus on semiconductor materials and molecular modeling.

You will source your own material, write prompts, and develop grading criteria while calibrating against failing models.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Materials Science PhD — AI Benchmark Architect
Materials Science PhD — AI Benchmark Architect

Mercor • New York (NY)

On-site
USD 83,000 - 124,000
AI Benchmark Scientist - Materials Science & Modeling (PhD)
AI Benchmark Scientist - Materials Science & Modeling (PhD)

Obsidian • New York (NY)

On-site
USD 83,000 - 138,000
6-Week Part-Time Materials Scientist for AI Benchmarks
6-Week Part-Time Materials Scientist for AI Benchmarks

Mercor • San Francisco (CA)

Remote
USD 150,000 - 190,000
Biochemist for AI Benchmark Tasks & Prompt Design
Biochemist for AI Benchmark Tasks & Prompt Design

Mercor • San Francisco (CA)

Remote
USD 96,000 - 138,000
Materials Science PhD - AI Evaluator
Materials Science PhD - AI Evaluator

Mercor • New York (NY)

On-site
USD 83,000 - 124,000
Materials Science PhD - AI Evaluator
Materials Science PhD - AI Evaluator

Obsidian • New York (NY)

On-site
USD 83,000 - 138,000
Quantum Chemistry AI Benchmark Scientist
Quantum Chemistry AI Benchmark Scientist

Obsidian • New York (NY)

On-site
USD 83,000 - 138,000
Materials Science Task Architect for AI Evaluation
Materials Science Task Architect for AI Evaluation

Mercor • San Francisco (CA)

On-site
USD 120,000 - 170,000
Quantum Chemistry AI Researcher & Code Architect
Quantum Chemistry AI Researcher & Code Architect

Obsidian • San Diego (CA)

On-site
USD 83,000 - 124,000
AI Benchmark Architect: Computational Mathematician
AI Benchmark Architect: Computational Mathematician

Mercor • San Francisco (CA)

Remote
USD 83,000 - 165,000