History & Political Science Benchmark Specialist

Mercor

San Francisco (CA)

On-site

USD 65,000 - 110,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Mercor is seeking experts in history and political science to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple-choice questions across core history and political science domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.

You will be assigned one of two task types: Question Authoring — Create original, challenging multiple-choice questions in your area

Qualifications

  • PhD or doctoral candidate in History, Political Science, International Relations, or closely related field.
  • Master's degree considered for exceptional depth in subdomain.
  • Strong command of historiographical methods, political theory, and comparative analysis.

Responsibilities

  • Author original history and political science questions that test deep conceptual understanding, not surface-level recall.
  • Ensure questions are unambiguous, self-contained, and precisely defined — all necessary information must be in the problem statement.
  • Rate each question's difficulty: Medium (intro undergraduate), Hard (advanced undergraduate), or Expert (post-graduate and above).
  • Provide 1 correct answer and 9 plausible but subtly incorrect alternatives that challenge expert-level solvers
  • Write step-by-step Chain-of-Thought solutions with clear, concise reasoning in markdown format
  • Supply 1–5 academic references per question from reputable sources (peer-reviewed journals, university repositories)
  • For verification tasks: flag issues with clarity, completeness, precision, or solvability and justify any edits made

Skills

Excellent written English
Historiographical methods
Comparative analysis

Education

PhD or doctoral candidate in History/Political Science/IR
Master's degree considered for exceptional depth

Job description

Mercor is seeking experts in history and political science to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple-choice questions across core history and political science domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.

You will be assigned one of two task types: Question Authoring — Create original, challenging multiple-choice questions in your area

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote History & Political Science Benchmark Specialist
Remote History & Political Science Benchmark Specialist

Mercor • United States

Remote
MXN 585,000 - 1,403,000
History & Political Science Assessment Specialist (Remote)
History & Political Science Assessment Specialist (Remote)

Mercor • San Francisco (CA)

On-site
USD 40,000 - 70,000
Remote work
Flexible schedule
History PhD - Benchmark Specialist
History PhD - Benchmark Specialist

Mercor • United States

Remote
MXN 585,000 - 1,403,000
Remote History & Political Science Question Architect
Remote History & Political Science Question Architect

Obsidian • San Francisco (CA)

On-site
USD 90,000 - 130,000
Remote AI Assessment Architect
Remote AI Assessment Architect

Mercor • New York (NY)

On-site
USD 55,000 - 124,000
Fully remote
Remote History & Political Science Assessment Specialist
Remote History & Political Science Assessment Specialist

Obsidian • San Francisco (CA)

On-site
USD 60,000 - 90,000
History PhD - Assessment Specialist - AI Trainer
History PhD - Assessment Specialist - AI Trainer

Mercor • San Francisco (CA)

On-site
USD 40,000 - 70,000
Remote work
Flexible schedule
Remote: Applied History & Political Science Specialist
Remote: Applied History & Political Science Specialist

United States Digital Space LLC • United States

Remote
USD 61,000 - 77,000
Remote Law Content Architect for AI Benchmarks
Remote Law Content Architect for AI Benchmarks

Mercor • New York (NY)

Remote
USD 80,000 - 120,000
Fully remote work
Asynchronous collaboration
Remote Mathematics Assessment Architect
Remote Mathematics Assessment Architect

Mercor • New York (NY)

Remote
USD 83,000 - 152,000