Quantum Computing AI Benchmark Scientist (PhD) — Part-Time

Mercor

New York (NY)

On-site

USD 83,000 - 124,000

Part time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Mercor is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code) for a new benchmark in scientific computing, collaborating with leading AI labs. You will author original, executable research problems that frontier models cannot solve.

Domains require depth in at least two subdomains with a coding focus. You'll use Python or similar, and work with GitHub and Docker to ensure runs pass automated quality checks.

Qualifications

  • PhD in physics, applied physics, or a closely related field.
  • Demonstrated depth in at least two subdomains: condensed matter, optics, quantum information/computing, computational physics, astrophysics, particle physics.
  • Proficiency in Python, R, or another relevant programming language for scientific computing.
  • Familiarity with Git/GitHub and running code in Docker with pull-request workflow and automated quality checks.

Responsibilities

  • Source your own material: a published paper, Kaggle dataset, an open-source repository, or a self-designed scenario.
  • Write scientific prompts based on the input.
  • Build the grading criteria that define a correct answer.
  • Calibrate against frontier models so tasks fail more often than they succeed.

Skills

Condensed matter physics
Quantum information/computing
Computational physics
Scientific problem design

Education

PhD in physics or closely related field

Tools

Git/GitHub
Docker
Jupyter

Job description

Mercor is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code) for a new benchmark in scientific computing, collaborating with leading AI labs. You will author original, executable research problems that frontier models cannot solve.

Domains require depth in at least two subdomains with a coding focus. You'll use Python or similar, and work with GitHub and Docker to ensure runs pass automated quality checks.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Quantum Computing Scientist for AI Benchmarks (PhD)
Quantum Computing Scientist for AI Benchmarks (PhD)

Mercor • San Francisco (CA)

On-site
USD 124,000 - 165,000
Quantum Chemistry AI Benchmark Engineer (PhD, Python)
Quantum Chemistry AI Benchmark Engineer (PhD, Python)

Mercor • San Diego (CA)

On-site
USD 9,919,000 - 14,878,000
Physics PhD - Quantum Computing Expert - AI Trainer
Physics PhD - Quantum Computing Expert - AI Trainer

Mercor • New York (NY)

On-site
USD 83,000 - 124,000
AI Evaluation Scientist: Math PhD for Frontier Benchmarks
AI Evaluation Scientist: Math PhD for Frontier Benchmarks

Mercor • San Francisco (CA)

On-site
USD 11,021,000 - 13,776,000
Physics PhD - Quantum Computing Expert
Physics PhD - Quantum Computing Expert

Mercor • San Francisco (CA)

On-site
USD 124,000 - 165,000
AI Evaluation Scientist - Math PhD (6-Week, Part-Time)
AI Evaluation Scientist - Math PhD (6-Week, Part-Time)

Mercor • Miami (FL)

On-site
USD 83,000 - 138,000
Bio PhD AI Benchmark Architect (Part-Time, 6-Week Project)
Bio PhD AI Benchmark Architect (Part-Time, 6-Week Project)

Mercor • San Diego (CA)

On-site
USD 69,000 - 103,000
Quantum Chemistry PhD - Coding Expert - AI Trainer
Quantum Chemistry PhD - Coding Expert - AI Trainer

Mercor • San Diego (CA)

On-site
USD 9,919,000 - 14,878,000
Mathematics PhD - AI Evaluation Expert
Mathematics PhD - AI Evaluation Expert

Mercor • San Francisco (CA)

On-site
USD 11,021,000 - 13,776,000
Mathematics PhD - AI Evaluation Expert - AI Trainer
Mathematics PhD - AI Evaluation Expert - AI Trainer

Mercor • Miami (FL)

On-site
USD 83,000 - 138,000