Engineering Manager, Safeguards & Interventions

Anthropic

San Francisco (CA)

On-site

USD 405,000 - 485,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Anthropic is seeking an Engineering Manager for Safeguards Interventions to lead the Interventions team responsible for what happens when a safety system fires across all Anthropic surfaces. You will own the roadmap, drive cross‑functional work with ML Infra, Research, Product, Policy, and Legal, and set the bar for safe, scalable interventions backed by data.

You should have a track record of leading engineering teams shipping production ML or safety‑enforcement systems, with comfort handling

Qualifications

  • Hands-on lead and grow a team of engineers; own roadmap, OKRs, and execution.
  • Drive cross-functional work with ML Infra, Research, Product, Policy, and Legal – and with cloud partners for 3P deployment.
  • Set the bar for when an intervention is good enough to ship – backed by measurement – and represent safety and product tradeoffs to leadership and external stakeholders.
  • Own production reliability for intervention and compliance systems: incident response, postmortems, SLOs, and the verification processes that prevent repeat incidents.
  • Have managed engineering teams shipping production ML or safety‑enforcement systems where the system’s decisions directly affected users.

Responsibilities

  • Lead and scale the Interventions team to deliver safe, reliable interventions.
  • Collaborate with ML Infra, Research, Product, Policy, and Legal to align safety with product goals.
  • Define measurement standards and governance for safety interventions and shipping decisions.
  • Ensure reliability and incident handling for safety-related production systems.
  • Communicate tradeoffs and outcomes to leadership and external stakeholders.

Skills

Team leadership
Safety systems
Cross-cloud experience
Stakeholder management

Job description

Anthropic is seeking an Engineering Manager for Safeguards Interventions to lead the Interventions team responsible for what happens when a safety system fires across all Anthropic surfaces. You will own the roadmap, drive cross‑functional work with ML Infra, Research, Product, Policy, and Legal, and set the bar for safe, scalable interventions backed by data.

You should have a track record of leading engineering teams shipping production ML or safety‑enforcement systems, with comfort handling

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Engineering Manager, Safety Interventions & ML Systems
Engineering Manager, Safety Interventions & ML Systems

Anthropic • San Francisco (CA)

Hybrid
USD 405,000 - 485,000
Engineering Manager, Safeguards Interventions San Francisco, CA
Engineering Manager, Safeguards Interventions San Francisco, CA

Anthropic • San Francisco (CA)

On-site
USD 405,000 - 485,000
Engineering Manager, Interventions & AI Safety
Engineering Manager, Interventions & AI Safety

United States Digital Space LLC • San Francisco (CA)

Hybrid
USD 405,000 - 485,000
Engineering Manager, Safeguards Interventions
Engineering Manager, Safeguards Interventions

Anthropic • San Francisco (CA)

Hybrid
USD 405,000 - 485,000
Engineering Manager, Safeguards Interventions
Engineering Manager, Safeguards Interventions

United States Digital Space LLC • San Francisco (CA)

Hybrid
USD 405,000 - 485,000
Engineering Manager, AI Safety Review Tooling
Engineering Manager, AI Safety Review Tooling

Anthropic • San Francisco (CA)

Hybrid
USD 405,000 - 485,000
Product Manager — AI Safety & Safeguards
Product Manager — AI Safety & Safeguards

Anthropic • San Francisco (CA)

Hybrid
USD 305,000 - 385,000
Tech Program Manager — Safeguards Infra & Evals
Tech Program Manager — Safeguards Infra & Evals

Anthropic • Seattle (WA)

On-site
USD 290,000 - 365,000
Staff Software Engineer – Safeguards Review Tooling
Staff Software Engineer – Safeguards Review Tooling

Menlo Ventures • San Francisco (CA)

Hybrid
USD 320,000 - 485,000
AI Safeguards Product Manager
AI Safeguards Product Manager

Visa Hunt • San Francisco (CA)

Hybrid
USD 305,000 - 385,000