Infrastructure Engineer, Interpretability & AI Safety

Anthropic Limited

New York (NY)

Hybrid

USD 320,000 - 485,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Anthropic is hiring a Software Engineer, Infrastructure, Interpretability to design, build, and own shared infrastructure for the Interpretability team. You will collaborate with agentic engineering, security, and platform teams to enable researchers to access frontier model weights securely and efficiently.

The role emphasizes security, privacy, data and compute management, and improving developer productivity across the research lifecycle, with a focus on scalable infrastructure in a rapidly

Qualifications

  • Proficient in at least one programming language (e.g., Python, Rust, Go, Java).
  • Experience building secure and scalable software infrastructure.

Responsibilities

  • Design, build, and own shared infrastructure for Interpretability - research environments, data systems, and compute tooling used daily by researchers.
  • Lead cross-team efforts across agentic engineering, security, compute, and storage platform teams to serve research needs.
  • Identify and resolve developer experience issues across the organization and help move interpretability methods from research to production-grade pipelines.

Skills

Python
Rust
Go
Java

Education

Bachelor's degree

Job description

Anthropic is hiring a Software Engineer, Infrastructure, Interpretability to design, build, and own shared infrastructure for the Interpretability team. You will collaborate with agentic engineering, security, and platform teams to enable researchers to access frontier model weights securely and efficiently.

The role emphasizes security, privacy, data and compute management, and improving developer productivity across the research lifecycle, with a focus on scalable infrastructure in a rapidly

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Infrastructure Engineer, Interpretability & AI Safety
Infrastructure Engineer, Interpretability & AI Safety

Anthropic • San Francisco (CA)

On-site
USD 190,000 - 270,000
Interpretability Research Engineer: Build Tools for Safe AI
Interpretability Research Engineer: Build Tools for Safe AI

Anthropic Limited • San Francisco (CA)

Hybrid
USD 315,000 - 560,000
Equity donation matching
Vacation and parental leave
Flexible working hours
+1
Software Engineer, Infrastructure, Interpretability
Software Engineer, Infrastructure, Interpretability

Anthropic • San Francisco (CA)

On-site
USD 190,000 - 270,000
Interpretability Engineer for AI Safety & Research
Interpretability Engineer for AI Safety & Research

Anthropic • San Francisco (CA)

Hybrid
USD 315,000 - 560,000
Remote Infrastructure Engineer for Interpretability & AI Safety
Remote Infrastructure Engineer for Interpretability & AI Safety

United States Digital Space LLC • San Francisco (CA), New York (NY)

On-site
USD 320,000 - 485,000
Software Engineer, Infrastructure, Interpretability
Software Engineer, Infrastructure, Interpretability

United States Digital Space LLC • San Francisco (CA), New York (NY)

On-site
USD 320,000 - 485,000
Staff Software Engineer, Safeguards Data Platforms
Staff Software Engineer, Safeguards Data Platforms

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 320,000 - 485,000
Competitive compensation
Benefits
Equity donation matching
+4
Research Engineer: AI Scientist Infra & Pipelines
Research Engineer: AI Scientist Infra & Pipelines

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Flexible hours
Generous vacation
Parental leave
+2
Mechanistic AI Interpretability Scientist
Mechanistic AI Interpretability Scientist

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Field Engineer, AI Interpretability & Deployment
Field Engineer, AI Interpretability & Deployment

Goodfire • San Francisco (CA)

On-site
USD 200,000 - 325,000
Market competitive salary
Equity
Competitive benefits