Research Manager, Interpretability — Lead Safe AI Teams

Anthropic

San Francisco (CA)

On-site

USD 340,000 - 425,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Competitive compensation
Generous vacation leave
Flexible working hours

Job summary

A leading AI research company based in San Francisco is seeking a Research Manager for Interpretability to lead a technical team. The role emphasizes project management, team coaching, and contributing to AI safety research. The ideal candidate will have 2-5 years of experience and a background in AI or machine learning, with strong communication skills. The salary range is competitive, offering $340,000—$425,000 annually.

Qualifications

  • Minimum 2–5 years of experience managing technical teams.
  • Background in machine learning, AI, or related technical field.
  • Strong project management skills including prioritization.

Responsibilities

  • Team support in achieving safety research goals.
  • Partner with research leads for project planning.
  • Coach and support team members.

Skills

Project management
People management
Machine learning
Communication

Education

Bachelor’s degree or equivalent experience

Job description

A leading AI research company based in San Francisco is seeking a Research Manager for Interpretability to lead a technical team. The role emphasizes project management, team coaching, and contributing to AI safety research. The ideal candidate will have 2-5 years of experience and a background in AI or machine learning, with strong communication skills. The salary range is competitive, offering $340,000—$425,000 annually.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Interpretability Research Manager — AI Safety Leadership
Interpretability Research Manager — AI Safety Leadership

Menlo Ventures • San Francisco (CA)

Hybrid
USD 340,000 - 425,000
Competitive salary
Generous vacation and parental leave
Flexible working hours
+1
Interpretability Engineer for AI Safety & Research
Interpretability Engineer for AI Safety & Research

Anthropic • San Francisco (CA)

Hybrid
USD 315,000 - 560,000
Mechanistic Interpretability Researcher for AI Safety
Mechanistic Interpretability Researcher for AI Safety

OpenAI • Los Angeles (CA)

On-site
USD 120,000 - 150,000
Senior AI Safety Leader — Strategy, Growth & Teams
Senior AI Safety Leader — Strategy, Growth & Teams

Safetytalent • San Francisco (CA)

On-site
USD 150,000 - 250,000
Research Engineer, Interpretability
Research Engineer, Interpretability

Anthropic • San Francisco (CA)

Hybrid
USD 315,000 - 560,000
AI Safety Research Manager: Lead Impactful Projects
AI Safety Research Manager: Lead Impactful Projects

Tais 2026 • United States

On-site
USD 130,000 - 200,000
Private medical insurance
Meals provided
Funding for professional development
+1
Research Lead, AI Safety & Impact Architect
Research Lead, AI Safety & Impact Architect

Aisafety • Berkeley (CA)

Hybrid
USD 170,000 - 270,000
Catered lunch and dinner
Visa sponsorship
Researcher, Alignment Interpretability
Researcher, Alignment Interpretability

OpenAI • San Francisco (CA)

On-site
USD 180,000 - 240,000
Research Engineer - Mechanistic Interpretability & AI Insight
Research Engineer - Mechanistic Interpretability & AI Insight

Acceler8 Talent • San Francisco (CA)

On-site
USD 140,000 - 190,000
Research Scientist — Safe, Steerable AI & LLMs
Research Scientist — Safe, Steerable AI & LLMs

Menlo Ventures • San Francisco (CA)

Hybrid
USD 340,000 - 425,000
Competitive compensation and benefits
Generous vacation and parental leave
Flexible working hours