Applied Math Benchmark Architect — Remote

Weekday AI

United States

Remote

USD 84,000 - 106,000

Part time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Fully remote

Job summary

Weekday AI is seeking expert mathematicians to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple-choice questions across core mathematics domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.

You will be assigned one of two task types: Question Authoring or Question Verification, with responsibilities spanning design, review, and documentation, all in a

Qualifications

  • PhD or doctoral candidate in Mathematics, Applied Mathematics, Statistics, or a related field.
  • Graduate-level mathematical proficiency and formal proof writing.
  • Experience with rigorous academic problem design or mathematical competition writing is a plus.
  • Excellent written English demonstrated in technical writing.

Responsibilities

  • Author original math questions that test deep conceptual understanding.
  • Ensure questions are unambiguous, self-contained, and precisely defined.
  • Rate each question’s difficulty: Medium, Hard, or Expert.
  • Provide 1 correct answer and 9 plausible distractors.
  • Write step-by-step Chain-of-Thought solutions in markdown format.
  • Supply 1–5 academic references per question from reputable sources.
  • For verification tasks, flag clarity, completeness, precision, or solvability issues and justify edits.

Skills

Excellent written English
Strong mathematical background
Question design
Proof-writing familiarity

Education

PhD in Mathematics
Doctoral candidate in Mathematics (or related field)
Master’s degree in Mathematics/Statistics (considered)

Job description

Weekday AI is seeking expert mathematicians to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple-choice questions across core mathematics domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.

You will be assigned one of two task types: Question Authoring or Question Verification, with responsibilities spanning design, review, and documentation, all in a

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Math Benchmark Architect — Remote
AI Math Benchmark Architect — Remote

Weekday 1 • United States

Remote
USD 84,000 - 106,000
Applied Mathematics Benchmark Specialist
Applied Mathematics Benchmark Specialist

Weekday 1 • United States

Remote
USD 84,000 - 106,000
Applied Mathematics Benchmark Specialist
Applied Mathematics Benchmark Specialist

Weekday AI • United States

Remote
USD 84,000 - 106,000
Fully remote
Remote AI Math Assessment Architect
Remote AI Math Assessment Architect

Obsidian • New York (NY)

On-site
USD 83,000 - 152,000
Remote Mathematics PhD — AI Assessment Question Architect
Remote Mathematics PhD — AI Assessment Question Architect

Obsidian • New York (NY)

On-site
USD 90,000 - 130,000
Remote Mathematics Assessment Architect
Remote Mathematics Assessment Architect

Mercor • New York (NY)

Remote
USD 83,000 - 152,000
Remote Mathematics Assessment Architect
Remote Mathematics Assessment Architect

Mercor • New York (NY)

Remote
USD 83,000 - 152,000
Remote Mathematics Researcher for AI Assessment Content
Remote Mathematics Researcher for AI Assessment Content

Obsidian • New York (NY)

Remote
USD 120,000 - 180,000
AI Benchmark CS Question Author (Remote Contractor)
AI Benchmark CS Question Author (Remote Contractor)

Weekday 1 • United States

Remote
USD 91,000 - 116,000
Fully remote
Remote CS Benchmark Architect for AI Evaluation
Remote CS Benchmark Architect for AI Evaluation

Weekday AI • United States

Remote
USD 91,000 - 116,000