Remote AI Benchmark Engineer — PhD-Level Question Author

Obsidian

Detroit (MI)

Vor Ort

USD 90.000 - 130.000

Teilzeit

14 Tage+
Bewerbungsgenerator

Eine Bewerbung wie gemacht für diesen Job — ein maßgeschneiderter Lebenslauf und ein Anschreiben, die genau zur Stellenanzeige passen.

Schaffe es an den ATS-Filtern vorbei

Zusammenfassung

Obsidian is seeking expert engineers to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple-choice questions across core engineering domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.

You will be assigned one of two task types: Question Authoring or Question Verification, both remote and asynchronous, with a commitment of 10+ hours per week.

Qualifikationen

  • PhD or doctoral candidate in Engineering or closely related field.
  • Excellent written English and ability to express complex ideas clearly.
  • Experience with academic assessment content is a plus.

Aufgaben

  • Author original engineering questions that test deep conceptual understanding, not surface-level recall.
  • Ensure questions are unambiguous, self-contained, and precisely defined — all necessary information must be in the problem statement.
  • Rate each question's difficulty: Medium (intro undergraduate), Hard (advanced undergraduate), or Expert (post-graduate and above).
  • Provide 1 correct answer and 9 plausible but subtly incorrect alternatives that challenge expert-level solvers.
  • Write step-by-step Chain-of-Thought solutions with clear, concise intermediate steps in markdown format.
  • Supply 1–5 academic references per question from reputable sources (peer-reviewed journals, university repositories).
  • For verification tasks: flag issues with clarity, completeness, precision, or solvability and justify any edits made.

Kenntnisse

Graduate-level engineering principles

Ausbildung

PhD
Master's degree (exceptional depth)

Jobbeschreibung

Obsidian is seeking expert engineers to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple-choice questions across core engineering domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.

You will be assigned one of two task types: Question Authoring or Question Verification, both remote and asynchronous, with a commitment of 10+ hours per week.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Remote AI Benchmark Engineer (PhD)
Remote AI Benchmark Engineer (PhD)

Obsidian • New York (NY)

Vor Ort
USD 90.000 - 130.000
Remote AI Engineering Assessment Author (PhD)
Remote AI Engineering Assessment Author (PhD)

Obsidian • San Francisco (CA)

Vor Ort
USD 83.000 - 124.000
Remote AI Assessment Architect (PhD) & Trainer
Remote AI Assessment Architect (PhD) & Trainer

Obsidian • Philadelphia

Vor Ort
USD 69.000 - 96.000
Remote Mathematics PhD — AI Assessment Question Architect
Remote Mathematics PhD — AI Assessment Question Architect

Obsidian • New York (NY)

Vor Ort
USD 90.000 - 130.000
Engineering PhD — AI Benchmark Question Architect
Engineering PhD — AI Benchmark Question Architect

Mercor • San Francisco (CA)

Vor Ort
USD 83.000 - 165.000
Fully remote
Asynchronous work
Remote AI Physics Question Designer (Applied Physics PhD)
Remote AI Physics Question Designer (Applied Physics PhD)

Obsidian • Miami (FL)

Vor Ort
USD 90.000 - 130.000
Remote AI Psychology Assessment Architect
Remote AI Psychology Assessment Architect

Obsidian • New York (NY)

Vor Ort
USD 55.000 - 110.000
Remote Physics Content Specialist - AI Benchmark & QA
Remote Physics Content Specialist - AI Benchmark & QA

Mercor • San Francisco (CA)

Remote
USD 55.000 - 90.000
Remote Law Content Author for AI Assessment & Benchmarks
Remote Law Content Author for AI Assessment & Benchmarks

Obsidian • New York (NY)

Remote
USD 60.000 - 90.000
Fully remote
Asynchronous work
10+ hours/week
Remote AI Physics Assessment Content Designer (PhD)
Remote AI Physics Assessment Content Designer (PhD)

Mercor • Philadelphia

Remote
USD 32.000 - 52.000
Fully remote