Mechanistic Interpretability Researcher

OpenAI

California (MO)

On-site

USD 180,000 - 320,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

OpenAI is seeking a researcher passionate about understanding deep networks, with a strong background in engineering, quantitative reasoning, and the research process. You will develop and carry out a research plan in mechanistic interpretability, in close collaboration with a highly motivated team.

You will help ensure future models remain safe as they grow, and you will play a critical role in OpenAI's mission to build and deploy safe AGI.

Qualifications

  • Ph.D. or equivalent research experience in computer science or ML.
  • Strong background in engineering, quantitative reasoning, and the research process.
  • 2+ years of research engineering experience.
  • Proficiency in Python or similar languages.

Responsibilities

  • Develop and publish research on techniques for understanding representations of deep networks.
  • Engineer infrastructure for studying model internals at scale.
  • Collaborate across teams on projects aligned with OpenAI's goals.
  • Guide research directions toward demonstrable usefulness and long-term scalability.

Skills

Python

Education

Ph.D. or equivalent research experience in CS/ML

Job description

OpenAI is seeking a researcher passionate about understanding deep networks, with a strong background in engineering, quantitative reasoning, and the research process. You will develop and carry out a research plan in mechanistic interpretability, in close collaboration with a highly motivated team.

You will help ensure future models remain safe as they grow, and you will play a critical role in OpenAI's mission to build and deploy safe AGI.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Mechanistic Interpretability Researcher
Mechanistic Interpretability Researcher

OpenAI • San Francisco (CA)

On-site
USD 180,000 - 240,000
Researcher, Interpretability
Researcher, Interpretability

OpenAI • Los Angeles (CA)

On-site
USD 120,000 - 150,000
Researcher, Interpretability
Researcher, Interpretability

OpenAI • California (MO)

On-site
USD 180,000 - 320,000
Mechanistic Interpretability Research Scientist
Mechanistic Interpretability Research Scientist

Anthropic • California (MO)

Hybrid
USD 350,000 - 850,000
Mechanistic Interpretability Researcher for AI Safety
Mechanistic Interpretability Researcher for AI Safety

OpenAI • Los Angeles (CA)

On-site
USD 120,000 - 150,000
Researcher, Alignment Interpretability
Researcher, Alignment Interpretability

OpenAI • San Francisco (CA)

On-site
USD 180,000 - 240,000
Mechanistic AI Interpretability Scientist
Mechanistic AI Interpretability Scientist

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Mechanistic Interpretability
Mechanistic Interpretability

Acceler8 Talent • San Francisco (CA)

On-site
USD 140,000 - 190,000
Research Engineer - Mechanistic Interpretability & AI Insight
Research Engineer - Mechanistic Interpretability & AI Insight

Acceler8 Talent • San Francisco (CA)

On-site
USD 140,000 - 190,000
Interpretability Research Engineer: Build Tools for Safe AI
Interpretability Research Engineer: Build Tools for Safe AI

Anthropic Limited • San Francisco (CA)

Hybrid
USD 315,000 - 560,000
Equity donation matching
Vacation and parental leave
Flexible working hours
+1