Conventional Weapons Safeguards Architect

EngineersOfAI

New York, Northern (NY, KY)

Hybrid

USD 120,000 - 180,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Anthropic is seeking a Safeguards Enforcement Analyst focused on Conventional Weapons. You will detect and mitigate attempts to misuse AI systems, build scalable enforcement workflows, and develop evaluations across technical policy areas.

The role requires handling potentially graphic or disturbing content in a controlled environment. Ideal candidates bring deep weapons systems expertise, policy enforcement experience, and strong data analysis skills, with the ability to collaborate across

Qualifications

  • Deep knowledge of weapons systems and ability to translate evidence into enforcement decisions.
  • Experience in policy enforcement, threat intelligence, or counterterrorism with exposure to harmful content or dangerous technology.
  • Experience standing up and scaling policy enforcement or content review workflows.
  • Proficiency with SQL or similar data analysis tools to draw insights from large datasets.

Responsibilities

  • Design and architect automated enforcement systems and scalable review workflows with high accuracy.
  • Develop and maintain evals to measure model performance on policy areas and surface regressions.
  • Collaborate with Engineering and Data Science to optimize detection and enforcement systems for policy violations.
  • Review flagged content to drive enforcement decisions and identify policy gaps.
  • Support policy design by providing feedback on enforcement ambiguities based on real scenarios.
  • Develop and maintain enforcement guidelines and reviewer documentation for consistent enforcement.
  • Stay updated on weapons trends, regulatory changes, and enforcement best practices.

Skills

Weapons systems expertise
Policy enforcement
Threat intelligence
Data analysis
Cross-functional communication
AI policy experience

Tools

SQL

Job description

Anthropic is seeking a Safeguards Enforcement Analyst focused on Conventional Weapons. You will detect and mitigate attempts to misuse AI systems, build scalable enforcement workflows, and develop evaluations across technical policy areas.

The role requires handling potentially graphic or disturbing content in a controlled environment. Ideal candidates bring deep weapons systems expertise, policy enforcement experience, and strong data analysis skills, with the ability to collaborate across

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Conventional Weapons Safeguards Analyst
Conventional Weapons Safeguards Analyst

Socket.dev • New York (NY)

Hybrid
USD 245,000 - 330,000
Conventional Weapons Safeguards Analyst — Remote
Conventional Weapons Safeguards Analyst — Remote

anthropic • New York (NY), San Francisco (CA), Washington

Hybrid
USD 245,000 - 330,000
Conventional Weapons Safeguards Analyst
Conventional Weapons Safeguards Analyst

United States Digital Space LLC • United States

Hybrid
USD 245,000 - 330,000
Safeguards Enforcement Analyst, Conventional Weapons
Safeguards Enforcement Analyst, Conventional Weapons

EngineersOfAI • New York (NY), Northern (KY)

Hybrid
USD 120,000 - 180,000
Policy Design Manager, Conventional Weapons
Policy Design Manager, Conventional Weapons

AI Chopping Block • New York (NY)

On-site
USD 245,000 - 285,000
Conventional Weapons Policy Architect
Conventional Weapons Policy Architect

Anthropic • New York (NY)

Hybrid
USD 245,000 - 285,000
Conventional Weapons Policy Lead
Conventional Weapons Policy Lead

Anthropic • Washington

On-site
USD 245,000 - 285,000
AI Integrity & Safeguards Enforcement Analyst
AI Integrity & Safeguards Enforcement Analyst

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000
Impact Safeguards Analyst — Violence & Extremism
Impact Safeguards Analyst — Violence & Extremism

Anthropic • New York (NY)

Hybrid
USD 285,000 - 330,000
Remote AI Safeguards & Cyber Enforcement Analyst
Remote AI Safeguards & Cyber Enforcement Analyst

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000
Equity donations
Flexible hours
Vacation policy
+2