Applied CS Benchmark Specialist - Remote Contract

Visa Hunt

United States

Remote

USD 91,000 - 116,000

Part time

8 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Mercor is seeking a contracted Applied Computer Science Benchmark Specialist to develop and evaluate advanced CS questions for AI research labs. Remote engagement requires ~10+ hours per week, with compensation available as an hourly rate.

The role emphasizes creating unambiguous questions, rating difficulty levels, and providing step-by-step explanations with references, while reviewing existing content for clarity and accuracy.

Qualifications

  • PhD in CS or closely related field.
  • Strong command of graduate-level CS theory, algorithms, systems design, and/or machine learning.
  • Excellent written English and ability to express complex ideas clearly and concisely.

Responsibilities

  • Author original computer science questions to test deep conceptual understanding.
  • Rate each question's difficulty as Medium, Hard, or Expert with correct and plausible distractors.
  • Write step-by-step Chain-of-Thought solutions with up to 5 references per question.
  • Review pre-written questions for accuracy and clarity and justify edits.
  • Work independently and asynchronously to meet deadlines.

Skills

PhD in CS/EE
Graduate CS theory
Algorithms
Systems design
Machine learning
Excellent English

Education

PhD or doctoral candidate
Master's degree considered

Job description

Mercor is seeking a contracted Applied Computer Science Benchmark Specialist to develop and evaluate advanced CS questions for AI research labs. Remote engagement requires ~10+ hours per week, with compensation available as an hourly rate.

The role emphasizes creating unambiguous questions, rating difficulty levels, and providing step-by-step explanations with references, while reviewing existing content for clarity and accuracy.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote CS Research Expert - Benchmark & ML Systems
Remote CS Research Expert - Benchmark & ML Systems

24-Mag Llc • New York (NY)

Remote
USD 76,000 - 103,000
Remote work
Flexible schedule
Competitive hourly rate
AI Benchmark Engineer — PhD (Remote & Flexible)
AI Benchmark Engineer — PhD (Remote & Flexible)

Mercor • New York (NY)

On-site
USD 55,000 - 110,000
Educational AI Benchmark Specialist (Remote, Part-Time)
Educational AI Benchmark Specialist (Remote, Part-Time)

Obsidian • Washington

On-site
USD 34,000 - 69,000
Remote Education AI Benchmark Specialist (Contractor)
Remote Education AI Benchmark Specialist (Contractor)

Mercor • New York (NY)

Remote
USD 14,000 - 20,000
Remote Part-Time Senior Engineering Assessment Consultant
Remote Part-Time Senior Engineering Assessment Consultant

24-Mag Llc • New York (NY)

Remote
USD 69,000 - 96,000
Remote Health AI Benchmark Architect
Remote Health AI Benchmark Architect

Mercor • United States

Remote
USD 55,000 - 110,000
Remote AI Benchmarking Consultant (Part-Time)
Remote AI Benchmarking Consultant (Part-Time)

Obsidian • San Francisco (CA)

On-site
USD 27,000 - 55,000
AI Education Benchmark Specialist (Remote)
AI Education Benchmark Specialist (Remote)

Obsidian • San Francisco (CA)

On-site
USD 27,000 - 46,000
AI Assessment Specialist — PhD Trainer (Remote)
AI Assessment Specialist — PhD Trainer (Remote)

Mercor • Philadelphia

On-site
USD 50,000 - 70,000
Remote AI Benchmark Test Engineer
Remote AI Benchmark Test Engineer

Mercor • New York (NY)

Remote
USD 85,000 - 120,000