Software Engineer — AI Safeguards & Oversight

SignalAI

New York (NY)

Hybrid

USD 320,000 - 485,000

Full time

13 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Anthropic is seeking software engineers to build safety and oversight mechanisms for AI systems on the Safeguards team. You will monitor models, prevent misuse, and ensure user well-being by detecting unwanted behaviors and enforcing policies.

Multiple Safeguards teams hire with flexible placement after interviews. You will work across sandboxed architectures and data pipelines to maintain safety, trust, and scale across Claude-based systems.

Qualifications

  • Bachelor's degree in Computer Science, Software Engineering or comparable experience.
  • Proficiency in Python and TypeScript.
  • Ability to work across the stack.
  • Strong communication skills and ability to explain complex technical concepts to non-technical stakeholders.

Responsibilities

  • Develop monitoring systems to detect unwanted behaviors from API partners and surface actions to analysts.
  • Build abuse detection mechanisms and infrastructure.
  • Surface abuse patterns to research teams to harden models at training.
  • Build reliable defenses for real-time safety improvements at scale.

Skills

Python
TypeScript
Distributed systems
API design
Communication

Education

Bachelor's degree in Computer Science or related field

Job description

Anthropic is seeking software engineers to build safety and oversight mechanisms for AI systems on the Safeguards team. You will monitor models, prevent misuse, and ensure user well-being by detecting unwanted behaviors and enforcing policies.

Multiple Safeguards teams hire with flexible placement after interviews. You will work across sandboxed architectures and data pipelines to maintain safety, trust, and scale across Claude-based systems.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Software Engineer - AI Safety & Safeguards
Staff Software Engineer - AI Safety & Safeguards

Menlo Ventures • New York (NY)

Hybrid
USD 320,000 - 485,000
Equity donation matching
Flexible hours
Office space
+1
Staff Software Engineer, Data Safeguards & Governance
Staff Software Engineer, Data Safeguards & Governance

Visa Hunt • New York (NY), San Francisco (CA)

Hybrid
USD 320,000 - 485,000
Staff Software Engineer, Safeguards Data Platforms
Staff Software Engineer, Safeguards Data Platforms

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 320,000 - 485,000
Competitive compensation
Benefits
Equity donation matching
+4
Software Engineer, Safeguards
Software Engineer, Safeguards

Anthropic • San Francisco (CA)

On-site
USD 320,000 - 485,000
Staff Data Platform Engineer, Safeguards
Staff Data Platform Engineer, Safeguards

Anthropic • San Francisco (CA)

Hybrid
USD 320,000 - 485,000
Equity donation matching
Generous vacation and parental leave
Flexible working hours
Staff Engineer, Safeguards Review Tooling & Automation
Staff Engineer, Safeguards Review Tooling & Automation

Anthropic • San Francisco (CA)

Hybrid
USD 320,000 - 485,000
Staff+ Software Engineer, Safeguards
Staff+ Software Engineer, Safeguards

Menlo Ventures • New York (NY)

Hybrid
USD 320,000 - 485,000
Equity donation matching
Flexible hours
Office space
+1
Hybrid Software Engineer — AI Safety Evaluations
Hybrid Software Engineer — AI Safety Evaluations

Anthropic • San Francisco (CA)

Hybrid
USD 320,000 - 485,000
Product Manager, Safeguards for Safe AI Platforms
Product Manager, Safeguards for Safe AI Platforms

Anthropic • San Francisco (CA)

Hybrid
USD 305,000 - 385,000
Staff Software Engineer — AI Safety Evaluation Systems
Staff Software Engineer — AI Safety Evaluation Systems

Menlo Ventures • New York (NY)

Hybrid
USD 320,000 - 485,000