Senior AI Safeguards Engineer — Distributed Systems

Anthropic

San Francisco (CA)

On-site

USD 190,000 - 270,000

Full time

47 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Health insurance
Parental leave
Flexible PTO
Equity
Wellness stipend
Relocation support
Commuter benefits
Education stipend
Home office stipend
Meals in office

Job summary

Anthropic is seeking an experienced software engineer to join the Safeguards team in San Francisco. You will build distributed, scalable systems that detect unwanted model behaviors and prevent disallowed use of models, while interfacing with analysts through internal dashboards.

You will lead multi-month infrastructure projects, deploy monitoring across clouds, and harden sandboxed runtimes for AI agents, ensuring data residency, auditability, and cost efficiency.

Qualifications

  • Experience with multi-cloud infrastructure and provider-agnostic designs.
  • Proficiency in Python and distributed systems.
  • Ability to explain complex technical concepts to non-technical stakeholders.
  • Experience with sandboxing and isolation technologies or systems language like Rust.
  • Bachelor’s degree in CS/SE or equivalent experience.
  • 8+ years in software engineering with production LLm-based agents experience.
  • Experience with regulated or sensitive data handling and data residency.

Responsibilities

  • Scope and lead multi-month infrastructure projects from design to production.
  • Develop monitoring systems to detect unwanted model behaviors and surface in dashboards.
  • Deploy monitoring across cloud environments with consistent pipelines and observability.
  • Design sandboxed runtime for AI agents with strict isolation and audit logging.
  • Manage cost and capacity for large volumes of agent work and meet SLOs.
  • Collaborate with engineers, researchers, security, privacy, and legal teams.

Skills

Python
Distributed Systems
Infrastructure as Code
Strong communication
Sandboxing / isolation tech
Rust
Cloud experience
LLM-based agents in production

Education

Bachelor’s degree in Computer Science or related field

Tools

Kubernetes
Terraform
CI/CD tooling
Containers

Job description

Anthropic is seeking an experienced software engineer to join the Safeguards team in San Francisco. You will build distributed, scalable systems that detect unwanted model behaviors and prevent disallowed use of models, while interfacing with analysts through internal dashboards.

You will lead multi-month infrastructure projects, deploy monitoring across clouds, and harden sandboxed runtimes for AI agents, ensuring data residency, auditability, and cost efficiency.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff Software Engineer — AI Safety & Distributed Systems
Staff Software Engineer — AI Safety & Distributed Systems

Anthropic • New York (NY)

Hybrid
USD 320,000 - 485,000
Staff Software Engineer, Distributed Systems for Safe AI
Staff Software Engineer, Distributed Systems for Safe AI

Anthropic • San Francisco (CA)

On-site
USD 170,000 - 250,000
Staff Distributed Systems Engineer (Safeguards)
Staff Distributed Systems Engineer (Safeguards)

EngineersOfAI • San Francisco (CA), Northern (KY)

On-site
USD 320,000 - 485,000
Staff Software Engineer, AI Safety & Abuse Detection
Staff Software Engineer, AI Safety & Abuse Detection

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 320,000 - 485,000
Staff Software Engineer, AI Safety & Safeguards
Staff Software Engineer, AI Safety & Safeguards

Jobs in JS • Seattle (WA)

Hybrid
USD 320,000 - 485,000
Staff Software Engineer - AI Safety & Safeguards
Staff Software Engineer - AI Safety & Safeguards

Menlo Ventures • New York (NY)

Hybrid
USD 320,000 - 485,000
Equity donation matching
Flexible hours
Office space
+1
Staff Software Engineer, AI Safety & Abuse Detection
Staff Software Engineer, AI Safety & Abuse Detection

Anthropic • San Francisco (CA)

Hybrid
USD 320,000 - 485,000
Equity donation matching
Generous vacation and parental leave
Flexible working hours
+1
Staff+ Software Engineer, Distributed Systems
Staff+ Software Engineer, Distributed Systems

Anthropic • San Francisco (CA)

On-site
USD 170,000 - 250,000
AI Integrity & Safeguards Enforcement Analyst
AI Integrity & Safeguards Enforcement Analyst

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000
Software Engineer, Safeguards
Software Engineer, Safeguards

Anthropic • San Francisco (CA)

On-site
USD 320,000 - 485,000