Cyber Harm Safeguards Analyst — Remote

Appliedmethods

San Francisco, Northern (CA, KY)

Hybrid

USD 285,000 - 330,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Anthropic is seeking a Safeguards Enforcement Analyst to review flagged content and enforcement actions across products, focusing on cyber threat mitigation and misuse of AI systems. You will work with cross-functional teams to surface model errors and policy gaps while maintaining high accuracy in enforcement decisions.

The role involves remote-friendly work, exposure to sensitive content, and collaboration with policy, engineering, and legal teams to uphold safe AI usage at scale.

Qualifications

  • Experience in cybersecurity, including offensive techniques, malware analysis, or vulnerability research.
  • Experience performing content review, abuse investigations, or policy enforcement at scale.
  • Proficiency in SQL and/or Python for data analysis and threat detection.
  • Experience with generative AI products and crafting effective prompts for review.

Responsibilities

  • Review flagged content and accounts to make accurate enforcement decisions.
  • Detect and mitigate misuse of AI systems for cyber operations.
  • Triage and escalate high-severity cases to stakeholders.
  • Provide feedback to policy design teams on gaps surfaced by enforcement.
  • Collaborate with Engineering and Data Science to surface detection issues and improve models.

Skills

Cybersecurity experience
Content review & enforcement
SQL
Python
Policy feedback & collaboration
Generative AI familiarity

Education

Bachelor's degree in a relevant field

Tools

SQL
Python

Job description

Anthropic is seeking a Safeguards Enforcement Analyst to review flagged content and enforcement actions across products, focusing on cyber threat mitigation and misuse of AI systems. You will work with cross-functional teams to surface model errors and policy gaps while maintaining high accuracy in enforcement decisions.

The role involves remote-friendly work, exposure to sensitive content, and collaboration with policy, engineering, and legal teams to uphold safe AI usage at scale.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote AI Safeguards & Cyber Enforcement Analyst
Remote AI Safeguards & Cyber Enforcement Analyst

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000
Equity donations
Flexible hours
Vacation policy
+2
Cyber Safeguards Enforcement Lead
Cyber Safeguards Enforcement Lead

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 280,000
AI Integrity & Safeguards Enforcement Analyst
AI Integrity & Safeguards Enforcement Analyst

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000
AI Safeguards Enforcement Specialist
AI Safeguards Enforcement Specialist

Anthropic • Washington

Hybrid
USD 285,000 - 330,000
Safeguards Enforcement Analyst, Cyber Harm
Safeguards Enforcement Analyst, Cyber Harm

Anthropic • Washington

Hybrid
USD 285,000 - 330,000
Safeguards Enforcement Analyst, Cyber Harm
Safeguards Enforcement Analyst, Cyber Harm

Anthropic • New York (NY)

On-site
USD 285,000 - 330,000
AI Safeguards Enforcement Analyst
AI Safeguards Enforcement Analyst

Anthropic • New York (NY)

Hybrid
USD 285,000 - 330,000
Safeguards Enforcement Analyst, Cyber Harm
Safeguards Enforcement Analyst, Cyber Harm

Doist • New York (NY), Washington, San Francisco (CA)

Hybrid
USD 285,000 - 330,000
Cyber Policy Analyst — AI Safeguards & Compliance
Cyber Policy Analyst — AI Safeguards & Compliance

Anthropic • San Francisco (CA)

On-site
USD 190,000 - 285,000
Safeguards Enforcement Analyst, Cyber Harm
Safeguards Enforcement Analyst, Cyber Harm

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000
Equity donations
Flexible hours
Vacation policy
+2