Applied Mathematics Benchmark Specialist

Weekday 1

United States

Remote

USD 84,000 - 106,000

Part time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Weekday 1 is seeking expert mathematicians to author and review high-quality academic assessment content for an AI research initiative. The role is fully remote with compensation of $61-$77 per hour.

You will write and verify rigorous MCQs across core mathematics domains and help establish gold-standard benchmarks for AI progression. Candidates will be assigned two task types: Question Authoring or Question Verification, with responsibilities including clarity, precision, and documentation of

Qualifications

  • PhD or doctoral candidate in Mathematics or closely related field.
  • Strong command of graduate-level mathematical concepts and formal proof writing.
  • Experience with rigorous academic problem design or mathematical competition writing is a strong plus.

Responsibilities

  • Author original math questions that test deep conceptual understanding and rate difficulty.
  • Ensure questions are unambiguous, self-contained, and precisely defined.
  • Rate each question's difficulty: Medium, Hard, or Expert.
  • Provide 1 correct answer and 9 plausible alternatives.
  • Write step-by-step Chain-of-Thought solutions with clear intermediate steps.
  • Supply 1-5 academic references per question.

Education

PhD or doctoral candidate in Mathematics
Master's degree considered for exceptional depth

Tools

Rigorous problem-design experience
Graduate-level proof-writing

Job description

This role is for one of our clients

Compensation: $61 - $77 per hour

We are seeking expert mathematicians to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple-choice questions across core mathematics domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.

You will be assigned one of two task types:

  • Question Authoring - Create original, challenging multiple-choice questions in your area of mathematical expertise, rate their difficulty, and submit them for review.
  • Question Verification - Review pre-written questions for accuracy, clarity, and rigor. Edit where needed, rate difficulty, and document any changes made.
Requirements
Mathematics Domains Covered

Signal Processing, Financial Mathematics & Actuarial Science, Mathematical Economics, Mathematical Modeling of Ecological & Biological Systems, Mathematical Programming & Combinatorial Optimization, Geomathematics & Climate Modeling.

Key Responsibilities
  • Author original math questions that test deep conceptual understanding, not surface-level recall
  • Ensure questions are unambiguous, self-contained, and precisely defined - all necessary information must be in the problem statement
  • Rate each question's difficulty: Medium (intro undergraduate), Hard (advanced undergraduate), or Expert (post-graduate and above)
  • Provide 1 correct answer and 9 plausible but subtly incorrect alternatives that challenge expert-level solvers
  • Write step-by-step Chain-of-Thought solutions with clear, concise intermediate steps in markdown format
  • Supply 1-5 academic references per question from reputable sources (peer-reviewed journals, university repositories)
  • For verification tasks: flag issues with clarity, completeness, precision, or solvability and justify any edits made
Ideal Qualifications
  • PhD or doctoral candidate in Mathematics, Applied Mathematics, Statistics, or a closely related field
  • Master's degree considered for candidates with exceptional depth in a specific subdomain
  • Strong command of graduate-level mathematical concepts and formal proof writing
  • Experience with rigorous academic problem design or mathematical competition writing is a strong plus
  • Excellent written English and ability to express complex ideas clearly and concisely
More About the Opportunity
  • Expected commitment: 10+ hours/week
  • Asynchronous, fully remote work

We consider all qualified applicants without regard to legally protected characteristics and provide reasonable accommodations upon request.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Mathematics PhD - Benchmark Specialist
Mathematics PhD - Benchmark Specialist

Mercor • New York (NY)

On-site
USD 83,000 - 179,000
Remote work
Flexible hours
Collaborative AI research environment
Applied Computer Science Benchmark Specialist
Applied Computer Science Benchmark Specialist

Weekday 1 • United States

Remote
USD 91,000 - 116,000
Fully remote
Remote Mathematics Researcher for AI Assessment Content
Remote Mathematics Researcher for AI Assessment Content

Obsidian • New York (NY)

Remote
USD 120,000 - 180,000
AI Math Benchmark Architect — Remote
AI Math Benchmark Architect — Remote

Weekday 1 • United States

Remote
USD 84,000 - 106,000
Remote Mathematics Assessment Architect
Remote Mathematics Assessment Architect

Mercor • New York (NY)

Remote
USD 83,000 - 152,000
Remote Mathematics PhD - AI Assessment Content Expert
Remote Mathematics PhD - AI Assessment Content Expert

Mercor • New York (NY)

On-site
USD 83,000 - 179,000
Remote work
Flexible hours
Collaborative AI research environment
Remote Mathematics PhD — AI Assessment Question Architect
Remote Mathematics PhD — AI Assessment Question Architect

Obsidian • New York (NY)

On-site
USD 90,000 - 130,000
AI Math Specialist: Question Author & Reviewer (Remote)
AI Math Specialist: Question Author & Reviewer (Remote)

Obsidian • San Francisco (CA)

On-site
USD 30,000 - 60,000
Mathematics Expert (LATAM & Europe)
Mathematics Expert (LATAM & Europe)

Anyone AI Inc. • United States

Remote
USD 45,000 - 65,000
Remote work
Contract / part-time
Immediate start
+1
Olympiad Mathematics Expert
Olympiad Mathematics Expert

Weekday 1 • United States

Remote
USD 74,000 - 128,000