Lead, Safeguards & Content Review Ops

Socket.dev

New York (NY)

Hybrid

USD 285,000 - 330,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Anthropic is seeking a Safeguards Enforcement Lead for the User Well-Being team. You will manage a team across policy areas, oversee content review workflows, and collaborate with Engineering and Legal to scale detection systems.

You will handle onboarding, QA, and escalation, while staying informed on evolving policy and legal frameworks. You will work with cross-functional partners to ensure accurate and consistent enforcement at scale, including reporting trends and coordinating external

Qualifications

  • Experience managing teams in the User Well-Being space.
  • Experience in trust & safety, content moderation operations, or policy enforcement with exposure to child safety, mental health, abuse and exploitation, and age assurance harm areas.
  • Experience managing or coordinating content review operations, including quality assurance and workflow management.
  • Experience standing up and scaling policy enforcement or content review workflows.
  • Proficiency in SQL and/or other data analysis tools to monitor workflow health, review queue metrics, and surface enforcement trends.
  • Experience identifying emerging risks and communicating findings to cross-functional stakeholders, such as Product, Policy, Engineering, and Legal teams.
  • Understanding of challenges involved in implementing product policies at scale in the content moderation space.

Responsibilities

  • Manage a team of individual contributors across multiple policy areas under the User Well-Being banner.
  • Serve as the primary point of contact for review partners conducting content review, including onboarding, training, quality assurance, and ongoing relationship management.
  • Design and improve enforcement workflows to scale effectively as volume grows, while maintaining high accuracy and consistency across review decisions.
  • Partner with Engineering and Data Science teams to optimize detection models and automated enforcement systems for User Well-Being policies.
  • Develop and maintain internal documentation, decision trees, and review guidelines that enable accurate and consistent enforcement at scale.
  • Keep up to date with emerging AI policy enforcement best practices, evolving legal frameworks, and developments in technology, and use these to inform our workflows.
  • Identify and report trends in misuse patterns to internal stakeholders, including Policy, Legal, and Trust & Safety leadership.
  • Coordinate reporting obligations to relevant external bodies (e.g., NCMEC) in accordance with applicable law and Anthropic policy.

Skills

Team management
Policy enforcement
Content review operations
Workflow scaling
SQL & data analysis
Cross-functional communication
Product policy implementation

Education

Bachelor’s degree

Tools

Python
Hash-matching tech

Job description

Anthropic is seeking a Safeguards Enforcement Lead for the User Well-Being team. You will manage a team across policy areas, oversee content review workflows, and collaborate with Engineering and Legal to scale detection systems.

You will handle onboarding, QA, and escalation, while staying informed on evolving policy and legal frameworks. You will work with cross-functional partners to ensure accurate and consistent enforcement at scale, including reporting trends and coordinating external

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Head of User Well-Being Safeguards & Enforcement
Head of User Well-Being Safeguards & Enforcement

EngineersOfAI • New York (NY), Northern (KY)

Hybrid
USD 180,000 - 260,000
Lead, Safeguards Enforcement — Remote & Travel‑Ready
Lead, Safeguards Enforcement — Remote & Travel‑Ready

anthropic • New York (NY), San Francisco (CA), Washington

Hybrid
USD 285,000 - 330,000
Competitive compensation
Equity donation matching
Generous vacation and parental leave
+2
Child Safety Safeguards Operations Analyst
Child Safety Safeguards Operations Analyst

Anthropic • New York (NY)

Hybrid
USD 245,000 - 285,000
Equity donation matching
Generous vacation
Parental leave
+2
Head of Safeguards & Content Enforcement
Head of Safeguards & Content Enforcement

United States Digital Space LLC • United States

Hybrid
USD 285,000 - 330,000
Safeguards Enforcement Lead, User Well-Being
Safeguards Enforcement Lead, User Well-Being

EngineersOfAI • New York (NY), Northern (KY)

Hybrid
USD 180,000 - 260,000
Cyber Safeguards Enforcement Lead
Cyber Safeguards Enforcement Lead

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 280,000
Remote-Friendly Safeguards & User Well-Being Analyst
Remote-Friendly Safeguards & User Well-Being Analyst

Anthropic • New York (NY), Washington, San Francisco (CA)

Hybrid
USD 245,000 - 285,000
Remote Operations Analyst, Child Safety Safeguards
Remote Operations Analyst, Child Safety Safeguards

Applied Methods Ltd • San Francisco (CA)

Remote
USD 245,000 - 285,000
Equity donation matching
Generous vacation
Parental leave
Staff Software Engineer, Safeguards Review Tooling
Staff Software Engineer, Safeguards Review Tooling

Anthropic • San Francisco (CA)

On-site
USD 320,000 - 485,000
Identity & Access Safeguards Analyst
Identity & Access Safeguards Analyst

Anthropic • New York (NY)

Hybrid
USD 285,000 - 330,000