AI CS Benchmark Engineer (Remote)

1000scholars

United States

Remote

USD 90,000 - 140,000

Part time

3 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Mercor is seeking expert computer scientists to author and review high-quality academic assessment content for an AI research initiative. You will create and verify rigorous CS questions, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.

Role includes writing original questions or verifying pre-written ones, rating difficulty, and providing references. Remote, asynchronous work with weekly payments and independent contractor terms.

Qualifications

  • PhD or doctoral candidate in CS, EE, or a closely related field.
  • Strong command of graduate-level CS theory, algorithms, systems design, and ML.
  • Publications or industry/research background is a plus.

Responsibilities

  • Author original CS questions that test deep conceptual understanding.
  • Rate difficulty: Medium, Hard, or Expert according to the rubric.
  • Provide 1 correct answer and 9 plausible distractors for each question.
  • Write step-by-step Chain-of-Thought solutions with intermediate steps.
  • Supply 1–5 academic references per question.

Skills

CS theory
Algorithms
Systems design
Machine learning

Education

PhD
Master's degree considered

Job description

Mercor is seeking expert computer scientists to author and review high-quality academic assessment content for an AI research initiative. You will create and verify rigorous CS questions, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.

Role includes writing original questions or verifying pre-written ones, rating difficulty, and providing references. Remote, asynchronous work with weekly payments and independent contractor terms.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Benchmark CS Question Author (Remote Contractor)
AI Benchmark CS Question Author (Remote Contractor)

Weekday 1 • United States

Remote
USD 91,000 - 116,000
Fully remote
Remote CS Benchmark Architect for AI Evaluation
Remote CS Benchmark Architect for AI Evaluation

Weekday AI • United States

Remote
USD 91,000 - 116,000
Remote Applied Physics Benchmark Specialist for AI Research
Remote Applied Physics Benchmark Specialist for AI Research

1000scholars • United States

Remote
USD 90,000 - 130,000
Engineering PhD — AI Benchmark Question Architect
Engineering PhD — AI Benchmark Question Architect

Mercor • San Francisco (CA)

On-site
USD 83,000 - 165,000
Fully remote
Asynchronous work
AI Math Benchmark Architect — Remote
AI Math Benchmark Architect — Remote

Weekday 1 • United States

Remote
USD 84,000 - 106,000
Remote AI Benchmark Engineer — PhD-Level Question Author
Remote AI Benchmark Engineer — PhD-Level Question Author

Obsidian • Detroit (MI)

On-site
USD 90,000 - 130,000
Applied Computer Science Benchmark Specialist
Applied Computer Science Benchmark Specialist

Weekday AI • United States

Remote
USD 91,000 - 116,000
[Contract] Applied Computer Science Benchmark Specialist
[Contract] Applied Computer Science Benchmark Specialist

1000scholars • United States

Remote
USD 90,000 - 140,000
Applied Computer Science Benchmark Specialist
Applied Computer Science Benchmark Specialist

Weekday 1 • United States

Remote
USD 91,000 - 116,000
Fully remote
Applied Math Benchmark Architect — Remote
Applied Math Benchmark Architect — Remote

Weekday AI • United States

Remote
USD 84,000 - 106,000
Fully remote