An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Mercor is assembling a panel of nuclear materials and safeguards experts to red-team frontier AI models. The goal is to test whether a model can correctly judge the misuse potential of a technical request - answering legitimate questions fully while refusing genuinely dangerous ones.
You will write challenging single-turn prompts at three levels (benign, dual-use, adversarial) and evaluate model responses against a defined policy standard, providing a reference answer with technical reasoning.
Mercor is assembling a panel of nuclear materials and safeguards experts to red-team frontier AI models. The goal is to test whether a model can correctly judge the misuse potential of a technical request - answering legitimate questions fully while refusing genuinely dangerous ones.
You will write challenging single-turn prompts at three levels (benign, dual-use, adversarial) and evaluate model responses against a defined policy standard, providing a reference answer with technical reasoning.