Safeguards Enforcement & Safety Evaluations Analyst

Anthropic

San Francisco (CA)

Hybrid

USD 230,000 - 270,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Anthropic is seeking a Safeguards Enforcement Analyst in San Francisco to ensure our models meet safety standards. This role is remote-friendly but requires occasional travel. You will collaborate with various teams to manage evaluations and drive improvements.

The ideal candidate has experience in trust and safety and is comfortable in fast-paced, ambiguous environments. Strong program management skills and a willingness to expand technical tools are essential.

Qualifications

  • Experience in trust and safety, content operations, or policy enforcement.
  • Ability to thrive in ambiguous, fast-moving environments.
  • Experience building processes from scratch.

Responsibilities

  • Support model launch readiness through evaluations.
  • Partner with policy and domain experts throughout evaluations.
  • Manage evaluation outcomes and drive mitigations when needed.

Skills

Trust and safety
Program management
Process building
Technical toolkit expansion
Data tools proficiency

Education

Bachelor's degree or equivalent

Tools

SQL
Dashboards
Spreadsheets

Job description

Anthropic is seeking a Safeguards Enforcement Analyst in San Francisco to ensure our models meet safety standards. This role is remote-friendly but requires occasional travel. You will collaborate with various teams to manage evaluations and drive improvements.

The ideal candidate has experience in trust and safety and is comfortable in fast-paced, ambiguous environments. Strong program management skills and a willingness to expand technical tools are essential.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Safety Evaluations Specialist, Safeguards
Safety Evaluations Specialist, Safeguards

Anthropic • United States

Remote
USD 120,000 - 180,000
Safeguards Enforcement Analyst, Safety Evaluations Remote-Friendly (Travel-Required) | San Fran[...]
Safeguards Enforcement Analyst, Safety Evaluations Remote-Friendly (Travel-Required) | San Fran[...]

Anthropic • San Francisco (CA)

Hybrid
USD 230,000 - 270,000
AI Integrity & Safeguards Enforcement Analyst
AI Integrity & Safeguards Enforcement Analyst

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000
Remote Operations Analyst, Child Safety Safeguards
Remote Operations Analyst, Child Safety Safeguards

Applied Methods Ltd • San Francisco (CA)

Remote
USD 245,000 - 285,000
Equity donation matching
Generous vacation
Parental leave
Staff Software Engineer, Safeguards Review Tooling
Staff Software Engineer, Safeguards Review Tooling

Anthropic • San Francisco (CA)

On-site
USD 320,000 - 485,000
Senior SRE: Safeguards ML Infra & Safe Deployments
Senior SRE: Safeguards ML Infra & Safe Deployments

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Cyber Harm Safeguards Analyst — Remote
Cyber Harm Safeguards Analyst — Remote

Appliedmethods • San Francisco (CA), Northern (KY)

Hybrid
USD 285,000 - 330,000
AI Safeguards Analyst – Violence & Extremism
AI Safeguards Analyst – Violence & Extremism

Anthropic • San Francisco (CA)

On-site
USD 285,000 - 330,000
Remote Chem & Explosives Safeguards Analyst
Remote Chem & Explosives Safeguards Analyst

Anthropic • New York (NY), Washington, San Francisco (CA)

Hybrid
USD 245,000 - 285,000
AI Safeguards Specialist — Contract (Remote Possible)
AI Safeguards Specialist — Contract (Remote Possible)

Employer.com • San Francisco (CA)

On-site
USD 165,000 - 235,000