A complete application in a minute — tailored resume and cover letter, ready to send.
Anthropic in San Francisco is seeking a software engineer for the Safeguards team to build safety and oversight mechanisms for our AI systems. You will monitor models, prevent misuse, and ensure user wellbeing across the stack.
This role focuses on detecting unwanted model behaviors, enforcing policies, and surfacing abuse patterns to researchers for hardening models at the training stage. You will contribute to scalable defenses and real-time safety improvements.
Anthropic in San Francisco is seeking a software engineer for the Safeguards team to build safety and oversight mechanisms for our AI systems. You will monitor models, prevent misuse, and ensure user wellbeing across the stack.
This role focuses on detecting unwanted model behaviors, enforcing policies, and surfacing abuse patterns to researchers for hardening models at the training stage. You will contribute to scalable defenses and real-time safety improvements.