Engineering Manager, Safeguards Review Tooling New San Francisco, CA

Alcides Fonseca

San Francisco (CA)

Hybrid

USD 405,000 - 485,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Alcides Fonseca is looking for an Engineering Manager in San Francisco to lead the Review Tooling team responsible for ensuring the deployment and functioning of their models and products.

This role demands a strong technical and management background to oversee the systems enabling safety investigations, and it involves collaboration across various disciplines.

Compensation is competitive, ranging from $405,000 to $485,000 annually with opportunities for career development and visa sponsorship.

Qualifications

  • 4+ years of management experience in software engineering.
  • 10+ years of industry software engineering experience.
  • Experience with investigating and enforcement tooling for safety.

Responsibilities

  • Lead and develop a team of engineers for review tooling.
  • Define vision and roadmap for review tooling platform.
  • Ensure effective collaboration with policy and legal teams.

Skills

Leadership
Software Engineering
Communication
Team Management
Technical Architecture
Cross-Functional Collaboration
Automation

Education

Bachelor’s degree or equivalent

Job description

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.

About the Role

The Safeguards team is responsible for ensuring our models and products are developed and deployed safely. We are seeking an Engineering Manager to lead our Review Tooling team, which builds the systems that humans—and increasingly Claude—use to investigate potential harms and take enforcement actions across Anthropic’s first‑party products and third‑party cloud platforms.

This foundational role owns the tools our safety investigators rely on, as well as the underlying platform. The platform includes analytics capabilities, privacy‑preserving primitives that keep review workflows compatible with our data‑retention commitments, and a sandbox environment for rapid development of new review interfaces. As model capabilities grow, you will drive the scale of review through automation, balancing human judgment and Claude assistance.

You’ll partner closely with policy, operations, data science, and legal teams to ensure our enforcement systems are effective, accurate, and trustworthy.

Key Responsibilities
  • Lead, grow, and develop a team of engineers building investigation, review, and enforcement tooling for both first‑party and third‑party platform surfaces
  • Define the vision and roadmap for our review tooling platform, including analytics, privacy‑compatible data access primitives, and a sandbox for rapidly developing new review interfaces
  • Drive the team’s strategy for scaling review through automation, enabling reviewers to use Claude effectively and building toward Claude‑assisted and Claude‑driven review workflows
  • Partner with policy, operations, legal, privacy, and data science stakeholders to translate enforcement and investigation needs into reliable, well‑designed systems
  • Ensure review tooling evolves alongside new privacy primitives and data‑retention commitments so reviewers can work without compromising user trust
  • Create clarity for the team and stakeholders in an ambiguous and evolving environment
  • Take an inclusive, equitable approach to hiring and coaching top technical talent, and maintain a high‑performing team
  • Contribute to engineering‑wide initiatives as a member of Anthropic’s engineering management community
  • Experience managing software engineering teams, including hiring, coaching, and developing engineers
  • Technical background in full‑stack or platform engineering, with the ability to engage deeply in architecture and design discussions
  • Experience shipping internal tools or platforms with demanding operational users, and a track record of improving their workflows measurably
  • Experience working cross‑functionally with non‑engineering partners such as operations, policy, or legal teams
  • Excellent communication skills, including the ability to explain technical tradeoffs to non‑technical stakeholders
  • Care about the societal impacts of AI and want your work to make powerful systems safer
Preferred Qualifications
  • 4+ years of management experience, 10+ years of industry software engineering experience
  • Experience building trust and safety, integrity, fraud, or abuse‑prevention tooling, or other systems supporting human review at scale
  • Experience integrating LLMs or agentic systems into operational workflows, or building human‑in‑the‑loop automation
  • Experience building developer platforms or extensible tooling frameworks that other teams build on top of
  • Experience supporting enforcement or moderation systems across multiple product surfaces, including enterprise or cloud platform contexts
Compensation

$405,000 – $485,000 USD annually.

Logistics

Minimum education: Bachelor’s degree or an equivalent combination of education, training, or experience.

Required field of study: Relevant coursework, training, or professional experience.

Minimum years of experience: Equivalent to internal job level requirements for the position.

Location‑based hybrid policy: Expect staff to be in one of our offices at least 25% of the time.

Visa sponsorship: We sponsor visas via an immigration lawyer for hired candidates.

EEO Statement

We do not discriminate on the basis of any protected group status under any applicable law. All qualified applicants are encouraged to apply. Anthropic is an equal‑opportunity employer.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Engineering Manager, Safeguards Review Tooling
Engineering Manager, Safeguards Review Tooling

Anthropic • San Francisco (CA)

Hybrid
USD 405,000 - 485,000
Staff+ Software Engineer, Safeguards Review Tooling New San Francisco, CA
Staff+ Software Engineer, Safeguards Review Tooling New San Francisco, CA

Anthropic • San Francisco (CA)

Hybrid
USD 320,000 - 485,000
Staff+ Software Engineer, Safeguards Review Tooling
Staff+ Software Engineer, Safeguards Review Tooling

Menlo Ventures • San Francisco (CA)

Hybrid
USD 320,000 - 485,000
Staff+ Software Engineer, Safeguards Human Review Tooling
Staff+ Software Engineer, Safeguards Human Review Tooling

Anthropic • San Francisco (CA)

On-site
USD 320,000 - 485,000
Safeguards Enforcement Analyst, Safety Evaluations Remote-Friendly (Travel-Required) | San Fran[...]
Safeguards Enforcement Analyst, Safety Evaluations Remote-Friendly (Travel-Required) | San Fran[...]

Anthropic • San Francisco (CA)

Hybrid
USD 230,000 - 270,000
Product Manager, Safeguards (Verticals)
Product Manager, Safeguards (Verticals)

Anthropic • San Francisco (CA)

Hybrid
USD 305,000 - 385,000
Staff+ Site Reliability Engineer, Safeguards ML Infra
Staff+ Site Reliability Engineer, Safeguards ML Infra

Anthropic • Seattle (WA), New York (NY), San Francisco (CA)

On-site
USD 405,000 - 485,000
Product Manager, Safeguards (Child Safety)
Product Manager, Safeguards (Child Safety)

Menlo Ventures • San Francisco (CA)

Hybrid
USD 305,000 - 385,000
Safeguards Enforcement Analyst, Cyber Harm
Safeguards Enforcement Analyst, Cyber Harm

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000
Equity donations
Flexible hours
Vacation policy
+2
Technical Program Manager, Reliability Engineering San Francisco, CA | New York City, NY | Seat[...]
Technical Program Manager, Reliability Engineering San Francisco, CA | New York City, NY | Seat[...]

Anthropic • New York (NY)

Hybrid
USD 290,000 - 365,000
Competitive compensation
Generous vacation and parental leave
Flexible working hours