Remote Math Benchmark Architect for AI Research

Mercor

United States

Remote

USD 90,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Mercor seeks expert mathematicians to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous MCQs across core mathematics domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.

You will be assigned one of two task types: Question Authoring or Question Verification, rate difficulty, provide 1 correct answer with 9 plausible alternatives, and cite 1–5 academic references

Qualifications

  • PhD or doctoral candidate in Mathematics or closely related field.
  • Strong command of graduate-level mathematical concepts and formal proof writing.
  • Excellent written English and ability to express complex ideas clearly.
  • Experience with rigorous academic problem design or competition writing is a plus.

Responsibilities

  • Author original math questions testing deep conceptual understanding.
  • Rate each question's difficulty: Medium, Hard, or Expert.
  • Provide 1 correct answer and 9 plausible distractors.
  • Write step-by-step Chain-of-Thought solutions in Markdown format.
  • Supply 1–5 academic references per question from reputable sources.
  • For verification tasks: flag issues with clarity, completeness, precision, or solvability and justify edits.

Skills

Advanced mathematical knowledge
Proof writing
Academic English writing
Editorial review

Education

PhD in Mathematics or related field
Master's degree considered

Job description

Mercor seeks expert mathematicians to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous MCQs across core mathematics domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.

You will be assigned one of two task types: Question Authoring or Question Verification, rate difficulty, provide 1 correct answer with 9 plausible alternatives, and cite 1–5 academic references

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote Mathematics Assessment Architect
Remote Mathematics Assessment Architect

Mercor • New York (NY)

Remote
USD 83,000 - 152,000
Remote Math Expert for AI Assessment Content
Remote Math Expert for AI Assessment Content

Mercor • New York (NY)

Remote
USD 55,000 - 110,000
Remote AI Math Researcher & Problem Designer
Remote AI Math Researcher & Problem Designer

Mercor • San Francisco (CA)

On-site
USD 83,000 - 124,000
Remote Mathematics PhD - AI Assessment Content Expert
Remote Mathematics PhD - AI Assessment Content Expert

Mercor • New York (NY)

On-site
USD 83,000 - 179,000
Remote work
Flexible hours
Collaborative AI research environment
Remote Mathematics Researcher for AI Assessment Content
Remote Mathematics Researcher for AI Assessment Content

Obsidian • New York (NY)

Remote
USD 120,000 - 180,000
Remote: Engineering Benchmark & Question Authoring
Remote: Engineering Benchmark & Question Authoring

Mercor • United States

Remote
USD 70,000 - 110,000
Fully remote
Remote AI Assessment Architect
Remote AI Assessment Architect

Mercor • New York (NY)

On-site
USD 55,000 - 124,000
Fully remote
Remote Mathematics PhD — AI Assessment Question Architect
Remote Mathematics PhD — AI Assessment Question Architect

Obsidian • New York (NY)

On-site
USD 90,000 - 130,000
Remote AI Math Assessment Architect
Remote AI Math Assessment Architect

Obsidian • New York (NY)

On-site
USD 83,000 - 152,000
Remote Applied Psychology Benchmark Architect
Remote Applied Psychology Benchmark Architect

Mercor • United States

Remote
USD 55,000 - 110,000