Staff Software Engineer, AI Safety & Safeguards

Jobs in JS

Seattle (WA)

Hybrid

USD 320,000 - 485,000

Full time

8 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Anthropic is seeking software engineers to build safety and oversight mechanisms for AI systems. As part of the Safeguards team, you will monitor models, prevent misuse, and enforce our terms of service while developing scalable defenses and detection systems.

The role focuses on detecting unwanted model behaviors, preventing disallowed use, and surfacing abuse patterns for internal review. A strong CS foundation and experience with Python/TypeScript are essential.

Qualifications

  • Bachelor’s degree in CS, SE or equivalent is required.
  • Proficiency in Python and TypeScript is required.
  • Ability to work across the stack and communicate complex concepts clearly.

Responsibilities

  • Develop monitoring systems to detect unwanted behaviors from API partners and surface results in dashboards.
  • Build abuse detection mechanisms and infrastructure.
  • Surface abuse patterns to research teams to harden models at training.
  • Develop multi-layered defenses for real-time safety improvements at scale.

Skills

Python
TypeScript
Across the stack
Strong communication

Education

Bachelor’s degree in Computer Science or equivalent

Job description

Anthropic is seeking software engineers to build safety and oversight mechanisms for AI systems. As part of the Safeguards team, you will monitor models, prevent misuse, and enforce our terms of service while developing scalable defenses and detection systems.

The role focuses on detecting unwanted model behaviors, preventing disallowed use, and surfacing abuse patterns for internal review. A strong CS foundation and experience with Python/TypeScript are essential.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff Software Engineer - AI Safety & Safeguards
Staff Software Engineer - AI Safety & Safeguards

Menlo Ventures • New York (NY)

Hybrid
USD 320,000 - 485,000
Equity donation matching
Flexible hours
Office space
+1
Staff Software Engineer, AI Safety & Abuse Detection
Staff Software Engineer, AI Safety & Abuse Detection

Anthropic • San Francisco (CA)

Hybrid
USD 320,000 - 485,000
Equity donation matching
Generous vacation and parental leave
Flexible working hours
+1
Staff Software Engineer, AI Safety & Abuse Detection
Staff Software Engineer, AI Safety & Abuse Detection

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 320,000 - 485,000
Staff Software Engineer — AI Safety & Distributed Systems
Staff Software Engineer — AI Safety & Distributed Systems

Anthropic • New York (NY)

Hybrid
USD 320,000 - 485,000
Staff Software Engineer, Distributed Systems for Safe AI
Staff Software Engineer, Distributed Systems for Safe AI

Anthropic • San Francisco (CA)

On-site
USD 170,000 - 250,000
Staff Software Engineer, AI Safety & Safeguards
Staff Software Engineer, AI Safety & Safeguards

Anthropic Limited • New York (NY), Northern (KY)

Hybrid
USD 320,000 - 485,000
Software Engineer, Safeguards
Software Engineer, Safeguards

Anthropic • San Francisco (CA)

On-site
USD 320,000 - 485,000
AI Safety & Oversight Engineer
AI Safety & Oversight Engineer

Anthropic • San Francisco (CA)

On-site
USD 320,000 - 485,000
Staff Distributed Systems Engineer (Safeguards)
Staff Distributed Systems Engineer (Safeguards)

EngineersOfAI • San Francisco (CA), Northern (KY)

On-site
USD 320,000 - 485,000
Staff+ Software Engineer, Distributed Systems
Staff+ Software Engineer, Distributed Systems

Anthropic • San Francisco (CA)

On-site
USD 170,000 - 250,000