AI Safeguards Enforcement Specialist

Anthropic

Washington (District of Columbia)

Hybrid

USD 285,000 - 330,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Anthropic is seeking an Enforcement Analyst to review content and enforce actions across our products, focusing on detecting and mitigating attempts to misuse AI systems for cyber operations. You will review flagged activity related to cyberattacks, malware development, and exploitation, with possible expansion to broader enforcement.

Safety is central to our mission, and you may need to respond to escalations during weekends and holidays.

Qualifications

  • Experience in cybersecurity, including knowledge of offensive techniques, exploit development, malware analysis, or vulnerability research.
  • Experience performing content review, abuse investigations, or policy enforcement at volume.
  • Proficiency in SQL and/or Python for data analysis and threat detection.
  • Experience identifying emerging risks and communicating findings to a diverse set of stakeholders, such as Product, Policy, Engineering, and Legal teams.
  • Experience working with generative AI products, including writing effective prompts for content review and enforcement.

Responsibilities

  • Review flagged content and accounts to make accurate, well‑documented enforcement decisions in line with our usage policies
  • Detect and mitigate potential misuse of AI systems to facilitate cyberattacks, malware creation, exploitation tooling, and related harmful cyber operations
  • Triage and elevate novel, ambiguous, or high‑severity cases to appropriate stakeholders
  • Provide detailed feedback to the Safeguards policy design team on policy gaps surfaced through real enforcement scenarios
  • Partner with Engineering and Data Science teams by surfacing detection model errors and quality signals from review to improve precision and recall
  • Maintain high accuracy and consistency standards across review queues
  • Keep up to date with emerging AI policy enforcement best practices, threat‑actor tactics, and the evolving cyber threat landscape, using these to inform enforcement decisions

Skills

Cybersecurity
SQL
Python
Content review
Generative AI

Education

Bachelor's degree or equivalent

Job description

Anthropic is seeking an Enforcement Analyst to review content and enforce actions across our products, focusing on detecting and mitigating attempts to misuse AI systems for cyber operations. You will review flagged activity related to cyberattacks, malware development, and exploitation, with possible expansion to broader enforcement.

Safety is central to our mission, and you may need to respond to escalations during weekends and holidays.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Safeguards Enforcement Analyst
AI Safeguards Enforcement Analyst

Anthropic • New York (NY)

Hybrid
USD 285,000 - 330,000
Remote AI Safeguards & Cyber Enforcement Analyst
Remote AI Safeguards & Cyber Enforcement Analyst

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000
Equity donations
Flexible hours
Vacation policy
+2
Cyber Safeguards Enforcement Lead
Cyber Safeguards Enforcement Lead

AISafety • Northern (KY), New York (NY)

Hybrid
USD 285,000 - 330,000
AI Integrity & Safeguards Enforcement Analyst
AI Integrity & Safeguards Enforcement Analyst

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000
Cyber Safeguards Enforcement Lead
Cyber Safeguards Enforcement Lead

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 280,000
AI Safeguards Enforcement Architect
AI Safeguards Enforcement Architect

Anthropic • New York (NY)

Hybrid
USD 285,000 - 330,000
Cyber Harm Safeguards Analyst — Remote
Cyber Harm Safeguards Analyst — Remote

Appliedmethods • San Francisco (CA), Northern (KY)

Hybrid
USD 285,000 - 330,000
AI Safeguards Enforcement Analyst
AI Safeguards Enforcement Analyst

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000
AI Safeguards Enforcement Lead
AI Safeguards Enforcement Lead

United States Digital Space LLC • United States

Hybrid
USD 285,000 - 330,000
AI Safeguards & Policy Enforcement Specialist
AI Safeguards & Policy Enforcement Specialist

The Sage Group • United States

On-site
USD 60,000 - 90,000