Engineering PhD — AI Benchmark Question Architect

Mercor

San Francisco (CA)

On-site

USD 83,000 - 165,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Fully remote
Asynchronous work

Job summary

Mercor is seeking expert engineers to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple-choice questions across core engineering domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.

This is a fully remote, asynchronous role with an expected commitment of 10+ hours per week.

Qualifications

  • PhD or doctoral candidate in Engineering or closely related field.
  • Master's degree considered for exceptional depth in a subdomain.
  • Strong command of graduate-level engineering principles and applied mathematics.
  • Excellent written English and ability to express complex ideas clearly.
  • Professional licensure (PE) or relevant industry experience is a plus.

Responsibilities

  • Author original engineering questions that test deep conceptual understanding.
  • Ensure questions are unambiguous, self-contained, and precisely defined.
  • Rate each question's difficulty: Medium, Hard, or Expert.
  • Provide 1 correct answer and 9 plausible distractors.
  • Write step-by-step Chain-of-Thought solutions with clear intermediate steps in markdown format.
  • Supply 1–5 academic references per question from reputable sources.
  • For verification tasks: flag issues with clarity, completeness, precision, or solvability and justify edits.

Skills

Question authoring
Question verification
Academic writing
English writing

Education

PhD in Engineering
Master's in Engineering

Job description

Mercor is seeking expert engineers to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple-choice questions across core engineering domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.

This is a fully remote, asynchronous role with an expected commitment of 10+ hours per week.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Benchmark Engineer — PhD (Remote & Flexible)
AI Benchmark Engineer — PhD (Remote & Flexible)

Mercor • New York (NY)

On-site
USD 55,000 - 110,000
Remote AI Benchmark Engineer — PhD-Level Question Author
Remote AI Benchmark Engineer — PhD-Level Question Author

Obsidian • Detroit (MI)

On-site
USD 90,000 - 130,000
Remote AI Benchmark Engineer (PhD)
Remote AI Benchmark Engineer (PhD)

Obsidian • New York (NY)

On-site
USD 90,000 - 130,000
Engineering PhD - Benchmark Specialist
Engineering PhD - Benchmark Specialist

Mercor • New York (NY)

On-site
USD 55,000 - 110,000
Remote AI Engineering Assessment Author (PhD)
Remote AI Engineering Assessment Author (PhD)

Obsidian • San Francisco (CA)

On-site
USD 83,000 - 124,000
Remote AI Assessment Architect
Remote AI Assessment Architect

Mercor • New York (NY)

On-site
USD 55,000 - 124,000
Fully remote
Remote Engineering PhD Assessment Specialist
Remote Engineering PhD Assessment Specialist

Mercor • New York (NY)

On-site
USD 48,000 - 96,000
AI Assessment Specialist — PhD Trainer (Remote)
AI Assessment Specialist — PhD Trainer (Remote)

Mercor • Philadelphia

On-site
USD 50,000 - 70,000
Economics & Finance AI Assessment Architect (Remote)
Economics & Finance AI Assessment Architect (Remote)

Mercor • New York (NY)

On-site
USD 83,000 - 138,000
Senior Mechanical Engineer – AI Benchmark Problems & Solutions
Senior Mechanical Engineer – AI Benchmark Problems & Solutions

Mercor • New York (NY)

Remote
USD 180,000 - 230,000