Interpretability Research Manager — AI Safety Leadership

Menlo Ventures

San Francisco (CA)

Hybrid

USD 340,000 - 425,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Competitive salary
Generous vacation and parental leave
Flexible working hours
Equity donation matching

Job summary

A leading AI research organization is seeking a Manager for its Interpretability team in San Francisco. The ideal candidate will have a strong background in managing technical teams and a passion for AI safety research. This role involves overseeing project execution, supporting team member development, and leading recruitment efforts. Competitive compensation of $340,000 – $425,000 annually is offered alongside flexible working hours and a collaborative office environment.

Qualifications

  • Minimum 2-5 years of experience managing technical teams.
  • Background in machine learning or AI.
  • Strong project management skills.

Responsibilities

  • Partner with a research lead on team direction and project execution.
  • Set and maintain execution standards and quality.
  • Drive the team's recruiting efforts and support team members.

Skills

People management
Project management
Technical understanding of AI/ML
Communication skills

Education

Bachelor's degree in a related field

Job description

A leading AI research organization is seeking a Manager for its Interpretability team in San Francisco. The ideal candidate will have a strong background in managing technical teams and a passion for AI safety research. This role involves overseeing project execution, supporting team member development, and leading recruitment efforts. Competitive compensation of $340,000 – $425,000 annually is offered alongside flexible working hours and a collaborative office environment.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Manager, Interpretability — Lead Safe AI Teams
Research Manager, Interpretability — Lead Safe AI Teams

Anthropic • San Francisco (CA)

On-site
USD 340,000 - 425,000
Interpretability Engineer for AI Safety & Research
Interpretability Engineer for AI Safety & Research

Anthropic • San Francisco (CA)

Hybrid
USD 315,000 - 560,000
Mechanistic Interpretability Researcher for AI Safety
Mechanistic Interpretability Researcher for AI Safety

OpenAI • Los Angeles (CA)

On-site
USD 120,000 - 150,000
Senior AI Safety Leader — Strategy, Growth & Teams
Senior AI Safety Leader — Strategy, Growth & Teams

Safetytalent • San Francisco (CA)

On-site
USD 150,000 - 250,000
AI Safety Research Manager: Lead Impactful Projects
AI Safety Research Manager: Lead Impactful Projects

Tais 2026 • United States

On-site
USD 130,000 - 200,000
Private medical insurance
Meals provided
Funding for professional development
+1
Operations Lead for AI Safety & Research Growth
Operations Lead for AI Safety & Research Growth

Heelsandtech • Berkeley (CA)

Hybrid
USD 140,000 - 200,000
Lunch and dinner for onsite employees
Visa sponsorship
Benefits provided
Research Lead, AI Safety & Impact Architect
Research Lead, AI Safety & Impact Architect

Aisafety • Berkeley (CA)

Hybrid
USD 170,000 - 270,000
Catered lunch and dinner
Visa sponsorship
Infrastructure Engineer, Interpretability & AI Safety
Infrastructure Engineer, Interpretability & AI Safety

Anthropic • San Francisco (CA)

On-site
USD 190,000 - 270,000
Research Engineer, Interpretability
Research Engineer, Interpretability

Anthropic • San Francisco (CA)

Hybrid
USD 315,000 - 560,000
Interpretability Research Engineer: Build Tools for Safe AI
Interpretability Research Engineer: Build Tools for Safe AI

Anthropic Limited • San Francisco (CA)

Hybrid
USD 315,000 - 560,000
Equity donation matching
Vacation and parental leave
Flexible working hours
+1