An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Mercor is assembling a panel of nuclear domain experts to red-team frontier AI models. The task is to craft challenging prompts across benign, dual-use, and adversarial levels, and to evaluate model responses against a policy standard.
You will write reference answers and the technical reasoning behind them. This is a writing-intensive role requiring safeguards experience and the ability to explain complex judgments to non-specialists.
Mercor is assembling a panel of nuclear domain experts to red-team frontier AI models. The goal is to test whether a model can correctly judge the misuse potential of a technical request — answering legitimate questions fully while refusing genuinely dangerous ones.
You will:
Write challenging single-turn prompts in your domain, labelled across three levels: benign, dual-use, and adversarial.
Evaluate the responses that come back against a defined policy standard, and judge whether each was handled correctly.
Write the reference answer; what a correct response looks like, and the technical reasoning for why.
We're looking for fuel-cycle depth plus some exposure to how material is controlled, accounted for, or diverted. Safeguards and nonproliferation experience is the strongest signal. The red-teaming task is to write prompts that pull a model past the boundary; the plausible research framing, the question that looks academic until you know what it's actually asking for — and then to explain in writing why the response was or wasn't acceptable. Examples of relevant backgrounds (Ideally w/Red-teaming experience):
This is writing-intensive work. Every judgment you make needs a written rationale that a non-specialist can follow. Prior technical writing, published research, or expert witness experience is a strong signal — please include a sample or link. You will also be reading and writing about misuse scenarios in your field for sustained periods. We brief experts on this in advance, and you can pause or step away at any point without penalty.
Your work here will not involve, and must not draw on, classified or export-controlled information, or anything covered by an NDA or prepublication review obligation. If you hold such obligations you may still be a good fit — tell us in your application and we will scope the work accordingly.