Get more replies from employers
Send a job-specific resume in minutes.
Anthropic is hiring software engineers to build safety and oversight mechanisms for AI systems. As a member of the Safeguards team, you will monitor models, prevent misuse, and ensure user wellbeing with scalable defenses and automated enforcement actions.
You will contribute to detecting unwanted model behaviors, surface abuse patterns to research teams, and strengthen safeguards across the stack using Python and TypeScript.
Anthropic is hiring software engineers to build safety and oversight mechanisms for AI systems. As a member of the Safeguards team, you will monitor models, prevent misuse, and ensure user wellbeing with scalable defenses and automated enforcement actions.
You will contribute to detecting unwanted model behaviors, surface abuse patterns to research teams, and strengthen safeguards across the stack using Python and TypeScript.