Investigative Web Researcher for AI Benchmarking

turing

United States

Remote

USD 83,000 - 124,000

Full time

4 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Turing is hiring for a contractor to build an evaluation benchmark for frontier AI browsing agents. Your work will involve designing challenging research problems that require an auditable evidence trail, starting from verifiable facts and reconstructing questions that make those facts hard to locate.

You'll deliver a structured output including a runnable set of clues across dates, people, places, organizations, works, events, records, and quantities, plus a validation trail showing search

Qualifications

  • Proven open-web research ability locating primary records from government and institutional databases.
  • Precision with sourcing—cite exact pages, tables, and sections.
  • Native or near-native written English.
  • Experience with LLM evaluation, red-teaming, or benchmark construction.
  • Experience in archival research, investigative journalism, or OSINT is a plus.

Responsibilities

  • Produce a natural-language research question with a stable, verifiable answer.
  • Create independent clues spanning dates, people, places, organisations, works, events, records, quantities.
  • Provide a validation record listing searches and results.

Skills

Open-web research ability
Primary source sourcing
English writing
Evidence trail documentation
LLM evaluation experience

Tools

JSON

Job description

Turing is hiring for a contractor to build an evaluation benchmark for frontier AI browsing agents. Your work will involve designing challenging research problems that require an auditable evidence trail, starting from verifiable facts and reconstructing questions that make those facts hard to locate.

You'll deliver a structured output including a runnable set of clues across dates, people, places, organizations, works, events, records, and quantities, plus a validation trail showing search

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Investigative Web Researcher for AI Benchmarking (Contract)
Investigative Web Researcher for AI Benchmarking (Contract)

Turing • San Francisco (CA)

On-site
USD 83,000 - 131,000
Web Researcher
Web Researcher

Turing • San Francisco (CA)

On-site
USD 83,000 - 131,000
Web Researcher
Web Researcher

turing • United States

Remote
USD 83,000 - 124,000
Remote Investigative Research Specialist
Remote Investigative Research Specialist

Turing Global India • United States

Remote
USD 83,000 - 152,000
Remote Web Research Specialist
Remote Web Research Specialist

Turing Global India • United States

Remote
USD 83,000 - 152,000
Remote Technical Writer - AI Benchmark Projects
Remote Technical Writer - AI Benchmark Projects

YO AI Labs • Town of Schroeppel (NY)

Remote
USD 55,000 - 117,000
Remote Technical Writer for AI Benchmark & Documentation
Remote Technical Writer for AI Benchmark & Documentation

YO AI Labs • San Francisco (CA)

Remote
USD 55,000 - 117,000
Remote Technical Writer for AI Benchmark Tasks
Remote Technical Writer for AI Benchmark Tasks

YO AI Labs • Town of Texas (WI)

Remote
USD 60,000 - 90,000
Remote AI Benchmark Technical Writer (Contractor)
Remote AI Benchmark Technical Writer (Contractor)

YO AI Labs • California (MO)

Remote
USD 55,000 - 110,000
Remote AI Benchmark Specialist — Scientific & Technical Prompts
Remote AI Benchmark Specialist — Scientific & Technical Prompts

Invisible Technologies Inc. • United States

Remote
USD 14,000 - 41,000