Infrastructure Engineer, Interpretability & AI Safety

Anthropic

San Francisco (CA)

On-site

USD 190,000 - 270,000

Full time

13 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Anthropic in San Francisco is hiring for an Interpretability Infrastructure Engineer to design and own the shared infrastructure for research environments, data systems, and compute tooling used by frontier AI researchers.

You will collaborate across agentic engineering, security, and platform teams to enable secure, scalable access to model weights and experiments, while improving developer experience and productivity for interpretability work.

Qualifications

  • Proficient in at least one programming language and productive with Python
  • Experience building and operating secure and scalable software infrastructure

Responsibilities

  • Design, build, and own shared infrastructure for Interpretability - research environments, data systems, and compute tooling that researchers rely on daily
  • Lead cross-team efforts with agentic engineering, security, compute, and storage platform teams
  • Discover and resolve major organization-wide developer experience issues
  • Help take interpretability methods from research code to dependable audit pipelines

Skills

Python
Rust
Go
Java

Tools

AWS
GCP
Kubernetes
Data warehouses

Job description

Anthropic in San Francisco is hiring for an Interpretability Infrastructure Engineer to design and own the shared infrastructure for research environments, data systems, and compute tooling used by frontier AI researchers.

You will collaborate across agentic engineering, security, and platform teams to enable secure, scalable access to model weights and experiments, while improving developer experience and productivity for interpretability work.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Infrastructure Engineer, Interpretability & AI Safety
Infrastructure Engineer, Interpretability & AI Safety

Anthropic Limited • New York (NY)

Hybrid
USD 320,000 - 485,000
Interpretability Engineer for AI Safety & Research
Interpretability Engineer for AI Safety & Research

Anthropic • San Francisco (CA)

Hybrid
USD 315,000 - 560,000
Interpretability Research Engineer: Build Tools for Safe AI
Interpretability Research Engineer: Build Tools for Safe AI

Anthropic Limited • San Francisco (CA)

Hybrid
USD 315,000 - 560,000
Equity donation matching
Vacation and parental leave
Flexible working hours
+1
Software Engineer, Infrastructure, Interpretability
Software Engineer, Infrastructure, Interpretability

Anthropic • San Francisco (CA)

On-site
USD 190,000 - 270,000
Remote Infrastructure Engineer for Interpretability & AI Safety
Remote Infrastructure Engineer for Interpretability & AI Safety

United States Digital Space LLC • San Francisco (CA), New York (NY)

On-site
USD 320,000 - 485,000
Field Engineer, AI Interpretability & Deployment
Field Engineer, AI Interpretability & Deployment

Goodfire • San Francisco (CA)

On-site
USD 200,000 - 325,000
Market competitive salary
Equity
Competitive benefits
Software Engineer, Infrastructure, Interpretability
Software Engineer, Infrastructure, Interpretability

United States Digital Space LLC • San Francisco (CA), New York (NY)

On-site
USD 320,000 - 485,000
Mechanistic AI Interpretability Scientist
Mechanistic AI Interpretability Scientist

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Research Engineer, Interpretability
Research Engineer, Interpretability

Anthropic • San Francisco (CA)

Hybrid
USD 315,000 - 560,000
Researcher, Interpretability
Researcher, Interpretability

OpenAI • Los Angeles (CA)

On-site
USD 120,000 - 150,000