Safeguards Enforcement Operations

Employer.com

San Francisco (CA)

On-site

USD 165,000 - 235,000

Full time

8 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Anthropic, a well-known AI safety company based in San Francisco, is seeking a Safeguards Enforcement Operations Specialist for a six-month on-site contract with potential for permanence. Remote consideration is available for the right candidate, and the role involves labeling data, reviewing user interactions, and enforcing policy guidelines to improve safety systems.

You may encounter explicit content in this role as part of safeguarding training and evaluation.

Qualifications

  • A quick learner with the ability to adapt rapidly to changing environments.
  • Demonstrate proactivity and strong independent work capabilities
  • Consistently meet deadlines and deliver high-quality work
  • Possess exceptional verbal and written communication skills
  • Thrive in collaborative team settings
  • Have a keen eye for detail and systematic approach to complex tasks
  • Show genuine interest in AI safety and ethical technology development

Responsibilities

  • Review customer signals, user exchanges, and code to identify potential policy violations
  • Enforce established Safeguards workflows and policy guidelines with precision and care
  • Collaborate with Strategy Analysts, Policy Managers, and Threat Intelligence Investigators to identify trends
  • Highlight potential policy gaps
  • Flag content for classifier refinement
  • Produce comprehensive weekly reports on operational workflows
  • Handle customer communications and appeals related to policy enforcement
  • Provide surge review support and content labeling when team needs arise
  • Triage and respond to cross-functional requests, offering timely assistance and insights

Skills

Quick learner
Proactive
Independent work
Deadline oriented
Strong communication
Team collaboration
Attention to detail
AI safety interest

Job description

Job Description

This job is a six-month contract working on-site for a well-known AI company. Remote work will be considered for the right candidate. The role has the potential to turn permenant.

As a Safeguards Enforcement Operations Specialist, you will be at the forefront of maintaining the integrity of our AI systems by labeling data, reviewing user interactions, enforcing policy guidelines, and collaborating across teams to continuously improve our safety mechanisms.

Important context for this role: In this position you may be exposed to and engage with explicit content spanning a range of topics, including those of a sexual, violent, or psychologically disturbing nature.

Responsibilities
  • Review customer signals, user exchanges, and code to identify potential policy violations
  • Enforce established Safeguards workflows and policy guidelines with precision and care
  • Collaborate with Strategy Analysts, Policy Managers, and Threat Intelligence Investigators to:
    • Identify emerging trends
    • Highlight potential policy gaps
    • Flag content for classifier refinement
    • Produce comprehensive weekly reports on operational workflows
    • Handle customer communications and appeals related to policy enforcement
    • Provide surge review support and content labeling when team needs arise
    • Triage and respond to cross-functional requests, offering timely assistance and insights
Qualifications
  • A quick learner with the ability to adapt rapidly to changing environments
  • Demonstrate proactivity and strong independent work capabilities
  • Consistently meet deadlines and deliver high-quality work
  • Possess exceptional verbal and written communication skills
  • Thrive in collaborative team settings
  • Have a keen eye for detail and systematic approach to complex tasks
  • Show genuine interest in AI safety and ethical technology development
Strong Candidates May Also Have
  • Experience in content moderation, Trust & Safety, Safeguards, or related fields
  • Familiarity with policy enforcement in digital platforms
  • Understanding of AI ethics and responsible technology deployment
  • Background in analyzing user behavior and identifying potential risks
Our Commitment to Diversity

We encourage applications from candidates of all backgrounds. Research shows that people from underrepresented groups are more likely to experience imposter syndrome. Anthropic is dedicated to building a diverse, inclusive, and equitable environment that represents a variety of perspectives and experiences. We believe diverse teams create more robust, thoughtful, and innovative solutions.

About Anthropic

Anthropic is a public benefit corporation headquartered in San Francisco, dedicated to ensuring that artificial intelligence systems are safe and beneficial to humanity. We are a team of researchers, engineers, and policy experts working together to develop responsible AI technologies.

Interested in helping shape the future of safe and ethical AI? We'd love to hear from you

Additional Information

All your information will be kept confidential according to EEO guidelines.

Compensation

$145-$145

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Safeguards Enforcement Operations
Safeguards Enforcement Operations

Job Mobz • San Francisco (CA)

On-site
USD 60,000 - 90,000
Safeguards Enforcement Analyst, Cyber Harm
Safeguards Enforcement Analyst, Cyber Harm

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000
Equity donations
Flexible hours
Vacation policy
+2
Safeguards Enforcement Analyst, Cyber Harm
Safeguards Enforcement Analyst, Cyber Harm

Anthropic • Washington

Hybrid
USD 285,000 - 330,000
Safeguards Enforcement Analyst, Cyber Harm
Safeguards Enforcement Analyst, Cyber Harm

Anthropic • New York (NY)

On-site
USD 285,000 - 330,000
Safeguards Enforcement Analyst, Access Controls & Identity
Safeguards Enforcement Analyst, Access Controls & Identity

Anthropic • New York (NY)

Hybrid
USD 285,000 - 330,000
Safeguards Enforcement Lead, Cyber Harms
Safeguards Enforcement Lead, Cyber Harms

Anthropic • Washington, San Francisco (CA), New York (NY)

On-site
USD 285,000 - 330,000
Safeguards Enforcement Analyst, Integrity & Authenticity
Safeguards Enforcement Analyst, Integrity & Authenticity

Anthropic • New York (NY)

Hybrid
USD 285,000 - 330,000
Staff+ Software Engineer, Safeguards Data
Staff+ Software Engineer, Safeguards Data

Visa Hunt • New York (NY), San Francisco (CA)

Hybrid
USD 320,000 - 485,000
Safeguards Enforcement Analyst, Conventional Weapons
Safeguards Enforcement Analyst, Conventional Weapons

anthropic • New York (NY), San Francisco (CA), Washington

Hybrid
USD 245,000 - 330,000
Staff+ Software Engineer, Safeguards Data
Staff+ Software Engineer, Safeguards Data

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 320,000 - 485,000
Competitive compensation
Benefits
Equity donation matching
+4