Nuclear Safeguards Red‑Team Analyst

Obsidian

New York (NY)

On-site

USD 90,000 - 130,000

Full time

5 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Mercor is assembling a panel of nuclear materials and safeguards experts to red-team frontier AI models. The goal is to test whether a model can correctly judge the misuse potential of a technical request — answering legitimate questions fully while refusing genuinely dangerous ones.

You will write prompts, evaluate responses, and draft reference answers with clear rationales, sitting on the line between routine professional questions and misuse scenarios.

Qualifications

  • Export control licensing, classification or enforcement — dual-use lists, catch-all controls, end-user screening
  • Treaty and agreement implementation — NPT, CTBT, safeguards agreements, bilateral cooperation agreements
  • Proliferation pathway and programme analysis at a laboratory, agency, ministry or research institute
  • Sanctions, procurement network or illicit trade analysis
  • Open-source and technical analysis of nuclear programmes

Responsibilities

  • Write challenging single-turn prompts in your domain, labelled across three levels: benign, dual-use, and adversarial
  • Evaluate the responses that come back against a defined policy standard, and judge whether each was handled correctly
  • Write the reference answer — what a correct response looks like, and the technical reasoning for why
  • This domain lives on the dual-use line more than most. The same control list that stops a shipment tells a procurement agent what to avoid declaring; the same pathway analysis that informs policy is a roadmap read from the other direction. A model that refuses a routine export-classification question is failing as badly as one that answers a question it shouldn't.

Job description

Mercor is assembling a panel of nuclear materials and safeguards experts to red-team frontier AI models. The goal is to test whether a model can correctly judge the misuse potential of a technical request — answering legitimate questions fully while refusing genuinely dangerous ones.

You will write prompts, evaluate responses, and draft reference answers with clear rationales, sitting on the line between routine professional questions and misuse scenarios.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Nuclear Safeguards Red-Team Researcher
Nuclear Safeguards Red-Team Researcher

Obsidian • San Francisco (CA)

On-site
USD 90,000 - 150,000
Nuclear Safeguards Red-Teaming Expert
Nuclear Safeguards Red-Teaming Expert

Mercor • San Francisco (CA)

On-site
USD 120,000 - 180,000
Nuclear Safeguards Red-Teaming Specialist
Nuclear Safeguards Red-Teaming Specialist

Obsidian • New York (NY)

On-site
USD 120,000 - 180,000
Nuclear Forensics Red Team Specialist
Nuclear Forensics Red Team Specialist

Mercor • New York (NY)

On-site
USD 90,000 - 130,000
Nuclear Engineer — AI Red-Teaming & Safeguards
Nuclear Engineer — AI Red-Teaming & Safeguards

Mercor • New York (NY)

Remote
USD 120,000 - 180,000
Nuclear Forensics Red-Teaming Specialist
Nuclear Forensics Red-Teaming Specialist

Obsidian • New York (NY)

On-site
USD 110,000 - 190,000
Nuclear Nonproliferation AI Safety Analyst
Nuclear Nonproliferation AI Safety Analyst

Mercor • New York (NY)

On-site
USD 90,000 - 120,000
Restaurant
Nuclear Safeguards Expert - Red Team - AI Trainer
Nuclear Safeguards Expert - Red Team - AI Trainer

Obsidian • New York (NY)

On-site
USD 120,000 - 180,000
Nuclear Medicine Physicist for AI Safety Red-Team
Nuclear Medicine Physicist for AI Safety Red-Team

Obsidian • New York (NY)

On-site
USD 120,000 - 180,000
Nuclear Security Red Team - AI Safety Evaluator
Nuclear Security Red Team - AI Safety Evaluator

Mercor • San Francisco (CA)

On-site
USD 120,000 - 170,000