A complete application in a minute — tailored resume and cover letter, ready to send.
Mercor in the United Kingdom seeks a researcher to design evaluation frameworks for frontier AI in drug discovery. You’ll craft problems that require genuine reasoning, assemble a data room from grounded sources, and define grading criteria focused on substance over style.
Commitment is 20–40 hours weekly; you’ll work in a browser studio with Claude Code and receive full reimbursement for Claude Max. You will defend rigorous solutions and own rework if reviewers request changes.
This is an evaluation framework for frontier AI models in drug research and development. Domain experts write realistic research problems wrapped around messy evidence worlds. The model works offline in a sandbox with open-source scientific tooling, receiving incomplete, indirect and sometimes misleading data, and must reach a defensible conclusion. The answer is present in the evidence but never stated.
Your job is to build a research problem so realistic and well-constructed that the model has to genuinely reason to solve it — it can't pattern-match or look the answer up.
You are the final authority on every scientific and difficulty question in your task.
Minimum 20 hours per week, up to 40.
Work happens in a browser-based studio plus Claude Code. A Claude Max subscription is required and is fully reimbursed.