Get more replies from employers
Send a job-specific resume in minutes.
Anthropic is seeking a Safeguards Analyst focused on Integrity & Authenticity to design automated enforcement workflows and review processes for detecting and mitigating misuse of AI systems.
You will partner with Engineering and Data Science to optimize detection models, review flagged content, and enforce policies addressing coordinated inauthentic behavior, election interference, and privacy harms.
Anthropic is seeking a Safeguards Analyst focused on Integrity & Authenticity to design automated enforcement workflows and review processes for detecting and mitigating misuse of AI systems.
You will partner with Engineering and Data Science to optimize detection models, review flagged content, and enforce policies addressing coordinated inauthentic behavior, election interference, and privacy harms.