Mechanistic Interpretability Researcher

OpenAI

San Francisco (CA)

On-site

USD 180,000 - 240,000

Full time

13 days ago
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

OpenAI is seeking a researcher passionate about understanding deep networks, with a strong background in engineering and quantitative reasoning. You will develop and carry out a research plan in mechanistic interpretability, collaborating with a motivated team to help ensure future models remain safe as they grow in capability.

You will publish research describing techniques for understanding representations, engineer scalable infrastructure, and guide directions toward measurable usefulness and

Qualifications

  • Ph.D. or equivalent research experience in CS/ML.
  • 2+ years of research engineering experience with Python or similar languages.
  • Interest in mechanistic interpretability and AI safety.

Responsibilities

  • Develop and publish research on techniques for understanding representations of deep networks.
  • Engineer infrastructure for studying model internals at scale.
  • Collaborate across teams on projects aligned with OpenAI's goals.
  • Guide research directions toward demonstrable usefulness and scalability.

Skills

Quantitative reasoning
Collaboration
Curiosity

Education

Ph.D. in Computer Science / ML or related field

Tools

Python

Job description

OpenAI is seeking a researcher passionate about understanding deep networks, with a strong background in engineering and quantitative reasoning. You will develop and carry out a research plan in mechanistic interpretability, collaborating with a motivated team to help ensure future models remain safe as they grow in capability.

You will publish research describing techniques for understanding representations, engineer scalable infrastructure, and guide directions toward measurable usefulness and

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Mechanistic Interpretability Researcher
Mechanistic Interpretability Researcher

OpenAI • California (MO)

On-site
USD 180,000 - 320,000
Mechanistic Interpretability Research Scientist
Mechanistic Interpretability Research Scientist

Anthropic • California (MO)

Hybrid
USD 350,000 - 850,000
Researcher, Interpretability
Researcher, Interpretability

OpenAI • Los Angeles (CA)

On-site
USD 120,000 - 150,000
Mechanistic AI Interpretability Scientist
Mechanistic AI Interpretability Scientist

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Researcher, Interpretability
Researcher, Interpretability

OpenAI • California (MO)

On-site
USD 180,000 - 320,000
Mechanistic Interpretability Researcher for AI Safety
Mechanistic Interpretability Researcher for AI Safety

OpenAI • Los Angeles (CA)

On-site
USD 120,000 - 150,000
Researcher, Alignment Interpretability
Researcher, Alignment Interpretability

OpenAI • San Francisco (CA)

On-site
USD 180,000 - 240,000
Mechanistic Interpretability
Mechanistic Interpretability

Acceler8 Talent • San Francisco (CA)

On-site
USD 140,000 - 190,000
Research Engineer - Mechanistic Interpretability & AI Insight
Research Engineer - Mechanistic Interpretability & AI Insight

Acceler8 Talent • San Francisco (CA)

On-site
USD 140,000 - 190,000
Interpretability Research Manager — Lead High-Impact ML
Interpretability Research Manager — Lead High-Impact ML

Anthropic • California (MO)

Hybrid
USD 350,000 - 500,000
Equity donation matching
Generous vacation
Parental leave
+2