AI Integrity & Safeguards Enforcement Analyst

Anthropic

San Francisco (CA)

Hybrid

USD 285,000 - 330,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Anthropic is seeking a Safeguards Enforcement Analyst to design scalable enforcement workflows and partner with engineering and data science teams to improve policy enforcement against AI-enabled influence operations and disinformation. The role involves reviewing content, guiding policy improvements, and staying current with evolving AI policy enforcement practices.

You will work in a remote-friendly, US-based setting with potential in SF, NY, or DC, collaborating with cross-functional teams

Qualifications

  • Experience in trust & safety, policy enforcement, threat intelligence, or a closely related field with a focus on influence operations, disinformation, coordinated inauthentic behavior, election integrity, or privacy and surveillance harms.
  • Experience standing up and scaling policy enforcement or content review workflows.
  • Proficiency in SQL and/or other data analysis tools to draw insights from large datasets.

Responsibilities

  • Design and architect automated enforcement systems and review workflows that scale effectively.
  • Partner with Engineering and Data Science teams to optimize detection models for policy violations and automated enforcement systems.
  • Review flagged content to drive enforcement and policy improvements.
  • Enforce usage policies with a focus on detecting and mitigating AI‑enabled influence operations, coordinated inauthentic behavior, election interference, and targeting, tracking, or surveillance of individuals and groups.

Skills

Trust & Safety
Policy Enforcement
Threat Intelligence
AI policy
Content Review
Cross-platform IR
Python
SQL
OSINT techniques
Data analysis

Education

Bachelor’s degree or equivalent

Tools

SQL
Python
OSINT tools

Job description

Anthropic is seeking a Safeguards Enforcement Analyst to design scalable enforcement workflows and partner with engineering and data science teams to improve policy enforcement against AI-enabled influence operations and disinformation. The role involves reviewing content, guiding policy improvements, and staying current with evolving AI policy enforcement practices.

You will work in a remote-friendly, US-based setting with potential in SF, NY, or DC, collaborating with cross-functional teams

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Safeguards Enforcement Architect
AI Safeguards Enforcement Architect

Anthropic • New York (NY)

Hybrid
USD 285,000 - 330,000
AI Safeguards & Integrity Enforcement Analyst
AI Safeguards & Integrity Enforcement Analyst

Anthropic • Washington

Hybrid
USD 285,000 - 330,000
AI Safeguards Enforcement Analyst
AI Safeguards Enforcement Analyst

Anthropic • New York (NY)

Hybrid
USD 285,000 - 330,000
Remote AI Safeguards & Cyber Enforcement Analyst
Remote AI Safeguards & Cyber Enforcement Analyst

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000
Equity donations
Flexible hours
Vacation policy
+2
AI Safeguards Enforcement Specialist
AI Safeguards Enforcement Specialist

Anthropic • Washington

Hybrid
USD 285,000 - 330,000
Cyber Harm Safeguards Analyst — Remote
Cyber Harm Safeguards Analyst — Remote

Appliedmethods • San Francisco (CA), Northern (KY)

Hybrid
USD 285,000 - 330,000
AI Safeguards Analyst – Violence & Extremism
AI Safeguards Analyst – Violence & Extremism

Anthropic • San Francisco (CA)

On-site
USD 285,000 - 330,000
Cyber Safeguards Enforcement Lead
Cyber Safeguards Enforcement Lead

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 280,000
Cyber Policy Analyst — AI Safeguards & Compliance
Cyber Policy Analyst — AI Safeguards & Compliance

Anthropic • San Francisco (CA)

On-site
USD 190,000 - 285,000
AI Safeguards Specialist — Contract (Remote Possible)
AI Safeguards Specialist — Contract (Remote Possible)

Employer.com • San Francisco (CA)

On-site
USD 165,000 - 235,000