Research Scientist, AI Interpretability - SF + Equity

Goodfire

San Francisco, Northern (CA, KY)

Hybrid

USD 200,000 - 400,000

Full time

5 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Goodfire in San Francisco is seeking a Research Scientist to advance new interpretability techniques for large AI models. You will work with a small, mission-driven team to conduct original research, build practical tools, and push the field forward.

You will contribute to interpretability mechanisms, moonshots, and applied research, with responsibilities including visualizing internal model structures, collaborating with engineering, and sharing work via publications and open-source.

Qualifications

  • PhD or equivalent experience in ML, CS, or a quantitative science.
  • Deep familiarity with large models and how they work.
  • Fluency in Python and ML frameworks such as PyTorch.
  • Strong writing and communication skills for explaining complex ideas.
  • Drive to move quickly and take ownership.

Responsibilities

  • Conduct original research in interpretability and related fields.
  • Prototype techniques to visualize and manipulate internal model structures.
  • Collaborate with engineering to turn research into production-ready tools.
  • Share your work through publications, demos, and open-source contributions.
  • Help define and evolve our research direction.

Skills

Python
PyTorch
Writing & communication
Ownership

Education

PhD or equivalent (ML/CS/quantitative science)

Job description

Goodfire in San Francisco is seeking a Research Scientist to advance new interpretability techniques for large AI models. You will work with a small, mission-driven team to conduct original research, build practical tools, and push the field forward.

You will contribute to interpretability mechanisms, moonshots, and applied research, with responsibilities including visualizing internal model structures, collaborating with engineering, and sharing work via publications and open-source.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Scientist
Research Scientist

Goodfire • San Francisco (CA), Northern (KY)

Hybrid
USD 200,000 - 400,000
Applied Interpretability Scientist for Real-World AI
Applied Interpretability Scientist for Real-World AI

Goodfire • San Francisco (CA)

Hybrid
USD 220,000 - 315,000
Equity
Competitive benefits
Hybrid work arrangement
Product Engineer: Build Interpretable AI Tools
Product Engineer: Build Interpretable AI Tools

Goodfire • San Francisco (CA)

On-site
USD 100,000 - 150,000
Product Engineer: AI Interpretability & Platform
Product Engineer: AI Interpretability & Platform

Goodfire • San Francisco (CA)

On-site
USD 100,000 - 140,000
Market competitive salary
Equity
Competitive benefits
Infrastructure Engineer, Interpretability & AI Safety
Infrastructure Engineer, Interpretability & AI Safety

Anthropic • San Francisco (CA)

On-site
USD 190,000 - 270,000
Research Engineer, Interpretability
Research Engineer, Interpretability

Anthropic • United States

Hybrid
USD 315,000 - 560,000
Remote work considered case-by-case
Visa sponsorship may be available
Office in San Francisco
Head of AI Training & Interpretability
Head of AI Training & Interpretability

Goodfire • San Francisco (CA)

On-site
USD 400,000 - 600,000
Equity
Competitive benefits
Researcher, Interpretability
Researcher, Interpretability

OpenAI • Los Angeles (CA)

On-site
USD 120,000 - 150,000
Senior Leader for Interpretable AI & Alignment
Senior Leader for Interpretable AI & Alignment

Association of Fundraising Professionals (AFP) Silicon Valley Chapter • San Francisco (CA)

On-site
USD 250,000 - 450,000
Mechanistic Interpretability
Mechanistic Interpretability

Acceler8 Talent • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 230,000