AI Safety Evaluations Analyst

Anthropic

New York (NY)

Hybrid

USD 230,000 - 270,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Competitive compensation
Equity donation matching
Generous vacation
Parental leave
Flexible working hours
Office space for collaboration

Job summary

Anthropic is seeking a Safeguards Enforcement Analyst focused on Safety Evaluations to join our cross-functional Safeguards team in a hybrid role. You will run evaluations, interpret results, and drive mitigations across policy, engineering, and product teams to ensure model safety and policy compliance.

Responsibilities include developing new evals, maintaining rigorous documentation, and building workflows that scale this work as Anthropic expands.

Qualifications

  • Experience in trust and safety, content operations, policy enforcement, or related operational role at a tech company.
  • Ability to navigate ambiguity and coordinate across teams to drive work to completion.
  • Experience building processes, workflows, or programs from scratch.

Responsibilities

  • Run evaluations, monitor results, and surface regressions to stakeholders.
  • Coordinate creation of new evals and ensure alignment with evolving policies and model capabilities.
  • Manage evaluation outcomes and drive mitigations as needed.
  • Design eval quality processes and frameworks to keep evaluations high-signal as models evolve.
  • Create product-specific evaluations as Anthropic's offerings expand.
  • Design tooling improvements for self-serve eval creation.

Skills

Trust and safety
Policy enforcement
Program management
Cross-functional collaboration
Data tools (SQL)

Education

Bachelor's degree

Tools

SQL
Dashboards

Job description

Anthropic is seeking a Safeguards Enforcement Analyst focused on Safety Evaluations to join our cross-functional Safeguards team in a hybrid role. You will run evaluations, interpret results, and drive mitigations across policy, engineering, and product teams to ensure model safety and policy compliance.

Responsibilities include developing new evals, maintaining rigorous documentation, and building workflows that scale this work as Anthropic expands.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Safeguards Enforcement & Safety Evaluations Analyst
Safeguards Enforcement & Safety Evaluations Analyst

Anthropic • San Francisco (CA)

Hybrid
USD 230,000 - 270,000
Safeguards Enforcement Analyst, Safety Evaluations Remote-Friendly (Travel-Required) | San Fran[...]
Safeguards Enforcement Analyst, Safety Evaluations Remote-Friendly (Travel-Required) | San Fran[...]

Anthropic • San Francisco (CA)

Remote
USD 230,000 - 270,000
Remote Operations Analyst, Child Safety Safeguards
Remote Operations Analyst, Child Safety Safeguards

Applied Methods Ltd • San Francisco (CA)

Remote
USD 245,000 - 285,000
Equity donation matching
Generous vacation
Parental leave
Child Safety Safeguards Operations Analyst
Child Safety Safeguards Operations Analyst

Anthropic • New York (NY)

Hybrid
USD 245,000 - 285,000
Equity donation matching
Generous vacation
Parental leave
+2
AI Safety & Cyber Evaluations Engineer
AI Safety & Cyber Evaluations Engineer

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 300,000 - 405,000
Identity & Access Safeguards Analyst
Identity & Access Safeguards Analyst

Anthropic • New York (NY)

Hybrid
USD 285,000 - 330,000
Biological Safety Research Manager
Biological Safety Research Manager

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 320,000 - 480,000
Staff Software Engineer, Safeguards Review Tooling
Staff Software Engineer, Safeguards Review Tooling

Anthropic • San Francisco (CA)

On-site
USD 320,000 - 485,000
AI Integrity & Safeguards Enforcement Analyst
AI Integrity & Safeguards Enforcement Analyst

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000
Identity & Access Enforcement Lead
Identity & Access Enforcement Lead

Anthropic • Washington

Hybrid
USD 285,000 - 330,000
Competitive compensation
Equity donation matching (optional)
Generous vacation
+3