Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
Turing in Austin, TX seeks a contractor to design research problems for frontier AI browsing agents, starting from a verifiable fact and crafting a question that is hard to locate. You will build an auditable evidence trail and produce a structured set of clues across dates, people, places, and organisations.
The role emphasizes open-web sourcing, precise citations, and experience with LLM evaluation or benchmarks. This is a contractor assignment for 8 weeks, 40 hours per week with PST overlap.
Turing is one of the worlds fastest-growing AI companies, accelerating the advancement and deployment of powerful AI systems.
Turing helps customers in two ways: Working with the worlds leading AI labs to advance frontier model capabilities in thinking, reasoning, coding, agentic behavior, multimodality, multilinguality, STEM and frontier knowledge; and leveraging that work to build real-world AI systems that solve mission-critical priorities for companies.
We are building an evaluation benchmark for frontier AI browsing agents. Your job is to design research problems that a state-of-the-art AI cannot solve, even with full web access and multiple attempts. This is not a subject matter expert role, nor is it a content-writing role. It is investigative research.
You will start from a verifiable fact, work backwards to construct a question that makes that fact extremely hard to locate, and then prove your work with a complete, auditable evidence trail.
Familiarity with JSON and structured data delivery formats