Remote: Engineering Benchmark & Question Authoring

Mercor

United States

Remote

USD 70,000 - 110,000

Full time

9 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Fully remote

Job summary

Mercor is seeking expert engineers to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple-choice questions across core engineering domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.

You will be assigned one of two task types: Question Authoring or Question Verification, with responsibilities spanning clarity, rigor, and documentation of edits.

Qualifications

  • PhD or doctoral candidate in Engineering or closely related field.
  • Master's degree considered for candidates with exceptional depth in a specific subdomain.
  • Excellent written English and ability to express complex ideas clearly and concisely.

Responsibilities

  • Author original engineering questions that test deep conceptual understanding, not surface-level recall.
  • Ensure questions are unambiguous, self-contained, and precisely defined — all necessary information must be in the problem statement.
  • Rate each question's difficulty: Medium (intro undergraduate), Hard (advanced undergraduate), or Expert (post-graduate and above).
  • Provide 1 correct answer and 9 plausible but subtly incorrect alternatives that challenge expert-level solvers.
  • Write step-by-step Chain-of-Thought solutions with clear, concise intermediate steps in markdown format.
  • Supply 1–5 academic references per question from reputable sources (peer-reviewed journals, university repositories).
  • For verification tasks: flag issues with clarity, completeness, precision, or solvability and justify any edits made.

Education

PhD in Engineering
Master's in Engineering (exceptional depth)

Job description

Mercor is seeking expert engineers to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple-choice questions across core engineering domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.

You will be assigned one of two task types: Question Authoring or Question Verification, with responsibilities spanning clarity, rigor, and documentation of edits.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote AI Assessment Architect
Remote AI Assessment Architect

Mercor • New York (NY)

On-site
USD 55,000 - 124,000
Fully remote
Remote AI Benchmark Engineer — PhD-Level Question Author
Remote AI Benchmark Engineer — PhD-Level Question Author

Obsidian • Detroit (MI)

On-site
USD 90,000 - 130,000
Remote Math Benchmark Architect for AI Research
Remote Math Benchmark Architect for AI Research

Mercor • United States

Remote
USD 90,000 - 150,000
Remote Applied Psychology Benchmark Architect
Remote Applied Psychology Benchmark Architect

Mercor • United States

Remote
USD 55,000 - 110,000
AI Benchmark Engineer — PhD (Remote & Flexible)
AI Benchmark Engineer — PhD (Remote & Flexible)

Mercor • New York (NY)

On-site
USD 55,000 - 110,000
Remote: Business & Commerce Benchmark Architect
Remote: Business & Commerce Benchmark Architect

Mercor • United States

Remote
USD 34,000 - 69,000
Remote AI Benchmark Engineer (PhD)
Remote AI Benchmark Engineer (PhD)

Obsidian • New York (NY)

On-site
USD 90,000 - 130,000
Remote Math Expert for AI Assessment Content
Remote Math Expert for AI Assessment Content

Mercor • New York (NY)

Remote
USD 55,000 - 110,000
Remote: Applied Chemistry Benchmark Specialist
Remote: Applied Chemistry Benchmark Specialist

Mercor • United States

Remote
USD 40,000 - 90,000
Fully remote
Remote Applied Physics Benchmark Architect
Remote Applied Physics Benchmark Architect

Mercor • United States

Remote
USD 34,000 - 83,000