Hebe dich für diese Rolle von der Masse ab — erstelle in etwa einer Minute einen maßgeschneiderten Lebenslauf und ein Anschreiben.
Obsidian seeks a researcher to craft rigorous, self-contained evaluation challenges for frontier AI models in drug discovery. You will design problems that force robust reasoning, assemble source-grounded data rooms, and write grading criteria focused on substance over method.
You will validate and revise iteratively within a browser-based studio environment, with Claude Code/Max tooling. Commitment ranges 20–40 hours weekly, requiring deep expertise in mechanistic enzymology or related fields,
This is an evaluation framework for frontier AI models in drug research and development. Domain experts write realistic research problems wrapped around messy evidence worlds. The model works offline in a sandbox with open-source scientific tooling, receiving incomplete, indirect and sometimes misleading data, and must reach a defensible conclusion. The answer is present in the evidence but never stated.
Your job is to build a research problem so realistic and well-constructed that the model has to genuinely reason to solve it — it can't pattern-match or look the answer up.
You are the final authority on every scientific and difficulty question in your task.
Minimum 20 hours per week, up to 40.
Work happens in a browser-based studio plus Claude Code. A Claude Max subscription is required and is fully reimbursed.