Product Manager, AI Safety & Safeguards

Anthropic

New York (NY)

Hybrid

USD 385,000 - 460,000

Full time

4 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Office space and amenities
Flexible working hours
Equity donation matching
Generous vacation and parental leave

Job summary

Anthropic is seeking a Product Manager for the Safeguards team to own ideation, design, development and deployment of safeguards systems and related UX. You will work with research and product teams to build detections, evals, interventions, and tools that mitigate deployment and user risks across Claude.ai, 1P API and cloud providers.

You will collaborate with policy, enforcement and engineering to define priorities, articulate tradeoffs, and drive safe, useful AI features in a fast-moving

Qualifications

  • Experience coordinating with policy, research, engineering and product teams.
  • Ability to design and measure safety-focused product features.
  • Strong written and verbal communication of complex concepts.

Responsibilities

  • Define safety-by-design features for frontier models across surfaces (Claude.ai, 1P API, cloud).
  • Write safety evals and communicate externally about safety.
  • Prioritize work with clear technical and business tradeoffs and MVP vs. ideal state.
  • Collaborate with policy, enforcement, research and engineering teams.
  • Lead development of metrics to assess risk, performance and gaps.

Skills

Tradeoff decisions
User understanding
Cross-functional leadership
Metrics design
Communication
Ambiguity management
Strategy development
Technical translation
Risk assessment

Education

Bachelor's degree

Job description

Anthropic is seeking a Product Manager for the Safeguards team to own ideation, design, development and deployment of safeguards systems and related UX. You will work with research and product teams to build detections, evals, interventions, and tools that mitigate deployment and user risks across Claude.ai, 1P API and cloud providers.

You will collaborate with policy, enforcement and engineering to define priorities, articulate tradeoffs, and drive safe, useful AI features in a fast-moving

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Product Manager, AI Safety & Safeguards
Product Manager, AI Safety & Safeguards

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 260,000
Product Manager, AI Safeguards
Product Manager, AI Safeguards

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 210,000
AI Safety & Safeguards Product Manager
AI Safety & Safeguards Product Manager

Alex Loftus • San Francisco (CA), Northern (KY)

Hybrid
USD 305,000 - 385,000
Product Manager - AI Safeguards & Risk Mitigation
Product Manager - AI Safeguards & Risk Mitigation

Anthropic Limited • San Francisco (CA), Northern (KY)

Hybrid
USD 385,000 - 460,000
AI Safety Product Manager — Safeguards
AI Safety Product Manager — Safeguards

Anthropic • New York (NY)

Hybrid
USD 385,000 - 460,000
Competitive compensation & benefits
Equity donation matching (optional)
Generous vacation & parental leave
+2
AI Safeguards Product Manager
AI Safeguards Product Manager

Anthropic • New York (NY)

Hybrid
USD 385,000 - 460,000
Product Manager, Safeguards (Account Integrity & Abuse)
Product Manager, Safeguards (Account Integrity & Abuse)

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 260,000
Product Manager, Safeguards (Generalist)
Product Manager, Safeguards (Generalist)

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 210,000
AI Safety & Safeguards Product Manager
AI Safety & Safeguards Product Manager

United States Digital Space LLC • San Francisco (CA), New York (NY)

On-site
USD 385,000 - 460,000
Head of AI Safety Policy & Harm Mitigation
Head of AI Safety Policy & Harm Mitigation

Anthropic • San Francisco (CA)

Hybrid
USD 330,000 - 395,000
Competitive compensation
Equity donation matching
Generous vacation and parental leave
+2