AI Safeguards Product Manager

Anthropic

New York (NY)

Hybrid

USD 385,000 - 460,000

Full time

3 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Anthropic is seeking a Product Manager for the Safeguards team to own ideation, design, development and deployment of Safeguards systems and relevant product UX to ensure we are advancing frontier models safely to users across various cloud platforms. You will work with research and product teams to develop detections, evals, interventions, and tools to measure and mitigate deployment and user risks.

We are looking for a product manager who is deeply committed to making AI safe and beneficial

Qualifications

  • Ability to make technical tradeoff decisions across policy experts, AI/ML research engineers and software engineering teams to design and build state of the art safety systems.
  • Strong user understanding of Safeguards concerns and how we provide the best solutions.
  • Demonstrated ability to build product and engineering strategy across multiple cross-functional teams in a rapidly changing space.
  • Demonstrated experience in designing and building metrics to evaluate risks, system performance and user impact.
  • Very strong ability to navigate and prioritize amidst rapidly changing product specs and to bring clarity across domains.
  • Evidence of exercising judgment and decision making in ambiguous situations.

Responsibilities

  • Determine how to build in safety by design upstream and leverage downstream defenses for Anthropic’s frontier models, AI products, customers on different surfaces - Claude.ai, 1P API, external Cloud providers.
  • Ability to write safety evals and communicate externally about safety.
  • Drive impact via ruthless prioritization by clearly defining problems, solution options forward, clarity on both business & technical tradeoffs and accordingly clear requirements toward MVP vs. ideal state.
  • Align & collaborate with policy, enforcement, research, engineering and cross functional stakeholders.
  • Understand the AI landscape and ecosystem to plan for mitigation of deployment risks of increasingly powerful models and determined adversaries.
  • Lead the development of metrics to understand the area, performance, blindspots to help inform future project planning.

Skills

Tradeoff decisions
Cross-functional leadership
Safety systems design
Metrics design
Stakeholder communication
Strategic thinking

Job description

Anthropic is seeking a Product Manager for the Safeguards team to own ideation, design, development and deployment of Safeguards systems and relevant product UX to ensure we are advancing frontier models safely to users across various cloud platforms. You will work with research and product teams to develop detections, evals, interventions, and tools to measure and mitigate deployment and user risks.

We are looking for a product manager who is deeply committed to making AI safe and beneficial

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Safety Product Manager — Safeguards
AI Safety Product Manager — Safeguards

Anthropic • New York (NY)

Hybrid
USD 385,000 - 460,000
Competitive compensation & benefits
Equity donation matching (optional)
Generous vacation & parental leave
+2
AI Safety & Safeguards Product Manager
AI Safety & Safeguards Product Manager

Alex Loftus • San Francisco (CA), Northern (KY)

Hybrid
USD 305,000 - 385,000
Product Manager - AI Safeguards & Risk Mitigation
Product Manager - AI Safeguards & Risk Mitigation

Anthropic Limited • San Francisco (CA), Northern (KY)

Hybrid
USD 385,000 - 460,000
Product Manager, AI Safety & Safeguards
Product Manager, AI Safety & Safeguards

Anthropic • New York (NY)

Hybrid
USD 385,000 - 460,000
Office space and amenities
Flexible working hours
Equity donation matching
+1
AI Safety & Safeguards Product Manager
AI Safety & Safeguards Product Manager

United States Digital Space LLC • San Francisco (CA), New York (NY)

On-site
USD 385,000 - 460,000
Product Manager, Safeguards (Cyber)
Product Manager, Safeguards (Cyber)

Alex Loftus • San Francisco (CA), Northern (KY)

Hybrid
USD 305,000 - 385,000
AI Safeguards Enforcement Architect
AI Safeguards Enforcement Architect

Anthropic • New York (NY)

Hybrid
USD 285,000 - 330,000
Tech Program Manager — Safeguards Infra & Evals
Tech Program Manager — Safeguards Infra & Evals

Anthropic • Seattle (WA)

On-site
USD 290,000 - 365,000
AI Safeguards Specialist — Contract (Remote Possible)
AI Safeguards Specialist — Contract (Remote Possible)

Employer.com • San Francisco (CA)

On-site
USD 165,000 - 235,000
Product Manager, Safeguards (Generalist)
Product Manager, Safeguards (Generalist)

Anthropic • New York (NY)

Hybrid
USD 385,000 - 460,000