AI Math Benchmark Architect — Remote

Weekday 1

United States

Remote

USD 84,000 - 106,000

Part time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Weekday 1 is seeking expert mathematicians to author and review high-quality academic assessment content for an AI research initiative. The role is fully remote with compensation of $61-$77 per hour.

You will write and verify rigorous MCQs across core mathematics domains and help establish gold-standard benchmarks for AI progression. Candidates will be assigned two task types: Question Authoring or Question Verification, with responsibilities including clarity, precision, and documentation of

Qualifications

  • PhD or doctoral candidate in Mathematics or closely related field.
  • Strong command of graduate-level mathematical concepts and formal proof writing.
  • Experience with rigorous academic problem design or mathematical competition writing is a strong plus.

Responsibilities

  • Author original math questions that test deep conceptual understanding and rate difficulty.
  • Ensure questions are unambiguous, self-contained, and precisely defined.
  • Rate each question's difficulty: Medium, Hard, or Expert.
  • Provide 1 correct answer and 9 plausible alternatives.
  • Write step-by-step Chain-of-Thought solutions with clear intermediate steps.
  • Supply 1-5 academic references per question.

Education

PhD or doctoral candidate in Mathematics
Master's degree considered for exceptional depth

Tools

Rigorous problem-design experience
Graduate-level proof-writing

Job description

Weekday 1 is seeking expert mathematicians to author and review high-quality academic assessment content for an AI research initiative. The role is fully remote with compensation of $61-$77 per hour.

You will write and verify rigorous MCQs across core mathematics domains and help establish gold-standard benchmarks for AI progression. Candidates will be assigned two task types: Question Authoring or Question Verification, with responsibilities including clarity, precision, and documentation of

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote AI Math Assessment Architect
Remote AI Math Assessment Architect

Obsidian • New York (NY)

On-site
USD 83,000 - 152,000
AI Benchmark Engineer — PhD (Remote & Flexible)
AI Benchmark Engineer — PhD (Remote & Flexible)

Mercor • New York (NY)

On-site
USD 55,000 - 110,000
Applied Mathematics Benchmark Specialist
Applied Mathematics Benchmark Specialist

Weekday 1 • United States

Remote
USD 84,000 - 106,000
AI Benchmark CS Question Author (Remote Contractor)
AI Benchmark CS Question Author (Remote Contractor)

Weekday 1 • United States

Remote
USD 91,000 - 116,000
Fully remote
Remote Mathematics PhD — AI Assessment Question Architect
Remote Mathematics PhD — AI Assessment Question Architect

Obsidian • New York (NY)

On-site
USD 90,000 - 130,000
AI Math Specialist: Question Author & Reviewer (Remote)
AI Math Specialist: Question Author & Reviewer (Remote)

Obsidian • San Francisco (CA)

On-site
USD 30,000 - 60,000
Remote Mathematics Researcher for AI Assessment Content
Remote Mathematics Researcher for AI Assessment Content

Obsidian • New York (NY)

Remote
USD 120,000 - 180,000
Remote Mathematics Assessment Architect
Remote Mathematics Assessment Architect

Mercor • New York (NY)

Remote
USD 83,000 - 152,000
Remote AI Math Researcher & Problem Designer
Remote AI Math Researcher & Problem Designer

Mercor • San Francisco (CA)

On-site
USD 83,000 - 124,000
Remote AI Math Assessment Architect (PhD)
Remote AI Math Assessment Architect (PhD)

Mercor • New York (NY)

On-site
USD 70,000 - 120,000