Safeguards Enforcement Analyst, Child Safety

Anthropic

New York (NY)

Hybrid

USD 245,000 - 285,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Equity donation matching
Generous vacation
Parental leave
Flexible hours
Office space

Job summary

Anthropic is seeking a Safeguards Enforcement Analyst to manage and scale child safety content review workflows, coordinating with engineering, policy, and legal teams to close enforcement gaps. You will oversee onboarding, training, QA, and ongoing relationships with review partners, ensuring accuracy and timely responses.

This role involves handling sensitive content and requires proactive design of scalable enforcement processes, documentation, and trend reporting to leadership.

Qualifications

  • Experience in trust & safety, content moderation operations, or policy enforcement with exposure to CSAM/CSEM.
  • Experience managing content review operations, including quality assurance and workflow management.
  • Experience standing up and scaling policy enforcement or content review workflows.
  • Proficiency in SQL and/or other data analysis tools to monitor workflow health, review queue metrics, and surface enforcement trends.
  • Experience communicating findings to cross‑functional stakeholders, such as Product, Policy, Engineering, and Legal teams.
  • Understanding of the challenges involved in implementing product policies at scale in the content moderation space.

Responsibilities

  • Own the day-to-day operational management of child safety content review workflows, including task routing, queue management, escalation handling, and SLA monitoring.
  • Serve as the primary point of contact for review partners conducting child safety content review, including onboarding, training, quality assurance, and ongoing relationship management.
  • Design and improve enforcement workflows to scale effectively as volume grows, while maintaining high accuracy and consistency across review decisions.
  • Partner with Engineering and Data Science teams to optimize detection models and automated enforcement systems for CSAM, CSEM, and related child safety policy violations.
  • Review novel or ambiguous flagged content to drive enforcement decisions and surface policy gaps to the Safeguards policy design team.
  • Develop and maintain internal documentation, decision trees, and review guidelines that enable accurate and consistent enforcement at scale.
  • Keep up to date with emerging AI policy enforcement best practices, evolving legal frameworks, and developments in child safety technology, and use these to inform our workflows.
  • Identify and report trends in misuse patterns to internal stakeholders, including Policy, Legal, and Trust & Safety leadership.
  • Coordinate reporting obligations to relevant external bodies (e.g., NCMEC) in accordance with applicable law and Anthropic policy.

Skills

Trust & Safety
Content moderation operations
Policy enforcement
SQL
Cross-functional communication
Policy implementation

Education

Bachelor’s degree or equivalent

Tools

Python
PhotoDNA
Perceptual hashing

Job description

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.

About the role

As a Safeguards Enforcement Analyst on the Child Safety team, you will be responsible for our child safety enforcement workflows, responsible for scaling, maintaining, and continuously improving the systems and processes we use to detect and respond to child sexual abuse material (CSAM) and child sexual exploitation material (CSEM) generated or facilitated through Anthropic's products.

This is a deeply operational role. You will serve as the central point of contact for those who conduct content review, managing day-to-day workflows, quality assurance, and escalation processes to ensure reviews are accurate, consistent, and conducted with appropriate support structures in place. You will also work closely with internal Engineering, Policy, and Legal teams to scale detection systems and close enforcement gaps as the threat landscape evolves.

This work is essential to Anthropic's mission. Child safety is one of our highest-priority harm areas, and the person in this role will have a direct and meaningful impact on protecting children from AI-facilitated exploitation and abuse.

Important context for this role: In this position you will regularly be exposed to and engage with explicit content of a sexual nature involving minors, as well as content that may be violent or psychologically disturbing. Anthropic takes the wellbeing of team members working in this area seriously and provides access to wellness resources and support. Candidates should carefully consider this aspect of the role before applying.

Key responsibilities
  • Own the day-to-day operational management of child safety content review workflows, including task routing, queue management, escalation handling, and SLA monitoring
  • Serve as the primary point of contact for review partners conducting child safety content review, including onboarding, training, quality assurance, and ongoing relationship management
  • Design and improve enforcement workflows to scale effectively as volume grows, while maintaining high accuracy and consistency across review decisions
  • Partner with Engineering and Data Science teams to optimize detection models and automated enforcement systems for CSAM, CSEM, and related child safety policy violations
  • Review novel or ambiguous flagged content to drive enforcement decisions and surface policy gaps to the Safeguards policy design team
  • Develop and maintain internal documentation, decision trees, and review guidelines that enable accurate and consistent enforcement at scale
  • Keep up to date with emerging AI policy enforcement best practices, evolving legal frameworks, and developments in child safety technology, and use these to inform our workflows
  • Identify and report trends in misuse patterns to internal stakeholders, including Policy, Legal, and Trust & Safety leadership
  • Coordinate reporting obligations to relevant external bodies (e.g., NCMEC) in accordance with applicable law and Anthropic policy
Minimum qualifications
  • Experience in trust & safety, content moderation operations, or policy enforcement with direct exposure to child safety, CSAM/CSEM, or related child protection harm areas
  • Experience managing or coordinating content review operations, including quality assurance and workflow management
  • Experience standing up and scaling policy enforcement or content review workflows
  • Proficiency in SQL and/or other data analysis tools to monitor workflow health, review queue metrics, and surface enforcement trends
  • Experience identifying emerging risks and communicating findings to cross‑functional stakeholders, such as Product, Policy, Engineering, and Legal teams
  • Understanding of the challenges involved in implementing product policies at scale in the content moderation space
Preferred qualifications
  • Deep subject matter expertise in child safety, child sexual exploitation and abuse (CSEA), or online child protection, including familiarity with CSAM/CSEM classification standards (e.g., COPINE, SAM scale)
  • Experience working with or reporting to NCMEC, IWF, or equivalent child safety reporting bodies
  • Familiarity with relevant legal and regulatory frameworks, including CSAM reporting obligations, KOSA, COPPA, or equivalent international frameworks
  • Experience working with generative AI products, including an understanding of how AI systems can be misused to generate or facilitate CSEA
  • Experience designing or evaluating trauma‑informed support structures and wellness protocols for content reviewers working with harmful material
  • Proficiency in Python for workflow automation or data analysis
  • Experience working with hash‑matching technologies (e.g., PhotoDNA, CSAI Match) or perceptual hashing tools used in CSAM detection
  • Familiarity with age assurance technologies and their role in child safety enforcement
  • Experience in a trust & safety role at a technology company, with an understanding of how platform policy intersects with child protection obligations
Compensation

Annual Salary: $245,000 – $285,000 USD

Logistics

Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience

Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience

Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position

Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices.

Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.

We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team.

Our safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings.

Benefits

Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Safeguards Enforcement Analyst, Child Safety
Safeguards Enforcement Analyst, Child Safety

Anthropic • San Francisco (CA)

Hybrid
USD 245,000 - 285,000
Safeguards Enforcement Lead, User Well-Being
Safeguards Enforcement Lead, User Well-Being

Socket.dev • New York (NY)

On-site
USD 285,000 - 330,000
Safeguards Enforcement Analyst, Cyber Harm
Safeguards Enforcement Analyst, Cyber Harm

Anthropic • San Francisco (CA)

On-site
USD 285,000 - 330,000
Equity donations
Flexible hours
Vacation policy
+2
1d Anthropic Safeguards Enforcement Analyst, Child Safety Remote-Friendly, United States; San F[...]
1d Anthropic Safeguards Enforcement Analyst, Child Safety Remote-Friendly, United States; San F[...]

Applied Methods Ltd • San Francisco (CA)

Remote
USD 245,000 - 285,000
Equity donation matching
Generous vacation
Parental leave
Safeguards Enforcement Analyst, Age-Appropriate Design
Safeguards Enforcement Analyst, Age-Appropriate Design

Anthropic • San Francisco (CA)

Hybrid
USD 245,000 - 285,000
Safeguards Enforcement Analyst, Cyber Harm
Safeguards Enforcement Analyst, Cyber Harm

Anthropic • New York (NY)

On-site
USD 285,000 - 330,000
Safeguards Enforcement Analyst, Integrity & Authenticity
Safeguards Enforcement Analyst, Integrity & Authenticity

Anthropic • San Francisco (CA)

On-site
USD 285,000 - 330,000
Safeguards Enforcement Analyst, Access Controls & Identity
Safeguards Enforcement Analyst, Access Controls & Identity

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000
Safeguards Enforcement Lead, User Well-Being
Safeguards Enforcement Lead, User Well-Being

EngineersOfAI • New York (NY), Northern (KY)

Hybrid
USD 180,000 - 260,000
Safeguards Enforcement Analyst, Cyber Harm
Safeguards Enforcement Analyst, Cyber Harm

Anthropic • Washington

Hybrid
USD 285,000 - 330,000