Safeguards Enforcement Lead, User Well-Being

United States Digital Space LLC

United States

Hybrid

USD 285,000 - 330,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

United States Digital Space LLC seeks a Safeguards Enforcement Lead on the User Well-Being team to manage enforcement workflows around child safety, mental health, abuse and exploitation, and age assurance. You will lead a reviewers' team and collaborate with Engineering, Policy, and Legal to scale detection and response.

This management role emphasizes process design, quality assurance, and cross-functional coordination in a high-stakes environment.

Qualifications

  • Experience managing teams in the User Well-Being space.
  • Experience in trust & safety, content moderation operations, or policy enforcement with exposure to child safety, mental health, abuse and exploitation, and age assurance harm areas.
  • Experience coordinating content review operations, including quality assurance and workflow management.

Responsibilities

  • Manage a team of individual contributors across multiple policy areas under the User Well-Being banner.
  • Serve as the primary point of contact for review partners conducting content review, including onboarding, training, quality assurance, and ongoing relationship management.
  • Design and improve enforcement workflows to scale effectively as volume grows, while maintaining high accuracy and consistency across review decisions.
  • Partner with Engineering and Data Science teams to optimize detection models and automated enforcement systems for User Well-Being policies.
  • Develop and maintain internal documentation, decision trees, and review guidelines that enable accurate and consistent enforcement at scale.
  • Keep up to date with emerging AI policy enforcement best practices, evolving legal frameworks, and developments in technology, and use these to inform our workflows.
  • Identify and report trends in misuse patterns to internal stakeholders, including Policy, Legal, and Trust & Safety leadership.
  • Coordinate reporting obligations to relevant external bodies (e.g., NCMEC) in accordance with applicable law and the company policy.

Skills

Team management
Content moderation
SQL analytics
Policy enforcement

Education

Bachelor's degree

Tools

SQL
Python

Job description

About the company

the company’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.

About the role

As a Safeguards Enforcement Lead on the User Well-Being team, you will be responsible for managing our child safety, mental health, abuse and exploitation, and age assurance enforcement workflows. This will include managing the team responsible for scaling, maintaining, and continuously improving the systems and processes we use to detect and respond to these harms.

This is a management role. You will serve as the central point of contact for those who conduct content review, managing day-to-day workflows, quality assurance, and escalation processes. You will also work closely with internal Engineering, Policy, and Legal teams to scale detection systems and close enforcement gaps as the threat landscape evolves.

*Important context for this role: In this position you will regularly be exposed to and engage with explicit content of a sexual nature involving minors, as well as content that may be violent or psychologically disturbing. the company takes the wellbeing of team members working in this area seriously and provides access to wellness resources and support. Candidates should carefully consider this aspect of the role before applying.*

Key responsibilities
  • Manage a team of individual contributors across multiple policy areas under the User Well-Being banner
  • Serve as the primary point of contact for review partners conducting content review, including onboarding, training, quality assurance, and ongoing relationship management
  • Design and improve enforcement workflows to scale effectively as volume grows, while maintaining high accuracy and consistency across review decisions
  • Partner with Engineering and Data Science teams to optimize detection models and automated enforcement systems for User Well-Being policies
  • Develop and maintain internal documentation, decision trees, and review guidelines that enable accurate and consistent enforcement at scale
  • Keep up to date with emerging AI policy enforcement best practices, evolving legal frameworks, and developments in technology, and use these to inform our workflows
  • Identify and report trends in misuse patterns to internal stakeholders, including Policy, Legal, and Trust & Safety leadership
  • Coordinate reporting obligations to relevant external bodies (e.g., NCMEC) in accordance with applicable law and the company policy
Minimum qualifications
  • Experience managing teams in the User Well-Being space
  • Experience in trust & safety, content moderation operations, or policy enforcement with direct exposure to child safety, mental health, abuse and exploitation, and age assurance harm areas
  • Experience managing or coordinating content review operations, including quality assurance and workflow management
  • Experience standing up and scaling policy enforcement or content review workflows
  • Proficiency in SQL and/or other data analysis tools to monitor workflow health, review queue metrics, and surface enforcement trends
  • Experience identifying emerging risks and communicating findings to cross-functional stakeholders, such as Product, Policy, Engineering, and Legal teams
  • Understanding of the challenges involved in implementing product policies at scale in the content moderation space
Preferred qualifications
  • Deep subject matter expertise in child safety, child sexual exploitation and abuse (CSEA), online child protection, mental wellness, and age assurance
  • Experience working with or reporting to NCMEC, IWF, or equivalent child safety reporting bodies
  • Familiarity with relevant legal and regulatory frameworks, including CSAM reporting obligations, KOSA, COPPA, or equivalent international frameworks
  • Experience working with generative AI products, including an understanding of how AI systems can be misused to generate or facilitate abusive content
  • Experience designing or evaluating trauma-informed support structures and wellness protocols for content reviewers working with harmful material
  • Proficiency in Python for workflow automation or data analysis
  • Experience working with hash-matching technologies (e.g., PhotoDNA, CSAI Match) or perceptual hashing tools used in CSAM detection
  • Familiarity with age assurance technologies and their role in child safety enforcement

The annual compensation range for this role is listed below.

For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role.

Annual Salary:

$285,000-$330,000 USD

Logistics

Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience

Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience

Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position

Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices.

Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.

Your safety matters to us.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Safeguards Enforcement Lead, User Well-Being
Safeguards Enforcement Lead, User Well-Being

EngineersOfAI • New York (NY), Northern (KY)

On-site
USD 180,000 - 260,000
Safeguards Enforcement Analyst, Age-Appropriate Design
Safeguards Enforcement Analyst, Age-Appropriate Design

Anthropic • New York (NY)

On-site
USD 245,000 - 285,000
Safeguards Enforcement Analyst, Child Safety
Safeguards Enforcement Analyst, Child Safety

Anthropic • New York (NY)

On-site
USD 245,000 - 285,000
Equity donation matching
Generous vacation
Parental leave
+2
Safeguards Enforcement Analyst, Cyber Harm
Safeguards Enforcement Analyst, Cyber Harm

Anthropic • New York (NY)

On-site
USD 285,000 - 330,000
Safeguards Enforcement Analyst, Age-Appropriate Design
Safeguards Enforcement Analyst, Age-Appropriate Design

Anthropic • San Francisco (CA)

Hybrid
USD 245,000 - 285,000
Safeguards Enforcement Analyst, Integrity & Authenticity
Safeguards Enforcement Analyst, Integrity & Authenticity

Anthropic • New York (NY)

On-site
USD 285,000 - 330,000
Safeguards Enforcement Analyst, User Well-being
Safeguards Enforcement Analyst, User Well-being

Anthropic • Washington

Hybrid
USD 245,000 - 285,000
Equity donation matching
Flexible working hours
Generous vacation and parental leave
+1
Safeguards Enforcement Lead, Cyber Harms
Safeguards Enforcement Lead, Cyber Harms

United States Digital Space LLC • United States

On-site
USD 285,000 - 330,000
Safeguards Enforcement Analyst, Conventional Weapons
Safeguards Enforcement Analyst, Conventional Weapons

United States Digital Space LLC • United States

On-site
USD 245,000 - 330,000
Safeguards Enforcement Analyst, Cyber Harm
Safeguards Enforcement Analyst, Cyber Harm

Anthropic • Washington

On-site
USD 285,000 - 330,000