Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
Mercor is assembling a panel of radiological safety experts to red-team frontier AI models. The goal is to test whether a model can correctly judge the misuse potential of a technical request — answering legitimate questions fully while refusing genuinely dangerous ones.
You will write prompts across benign, dual-use, and adversarial levels, evaluate responses against a defined policy standard, and craft the reference answer with technical reasoning for why.
Mercor is assembling a panel of radiological safety experts to red-team frontier AI models. The goal is to test whether a model can correctly judge the misuse potential of a technical request — answering legitimate questions fully while refusing genuinely dangerous ones.
You will write prompts across benign, dual-use, and adversarial levels, evaluate responses against a defined policy standard, and craft the reference answer with technical reasoning for why.