Get more replies from employers
Send a job-specific resume in minutes.
Anthropic is looking for software engineers in San Francisco to help build safety and oversight mechanisms for AI systems. You will monitor models, prevent misuse, and ensure user well-being by developing systems for detecting unwanted model behaviors.
Candidates are expected to have a Bachelor’s degree in Computer Science or Software Engineering and proficiency in Python and Typescript. The annual compensation range is $320,000 - $485,000.
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.
We are looking for software engineers to help build safety and oversight mechanisms for our AI systems. As a software engineer on the Safeguards team, you will work to monitor models, prevent misuse, and ensure user well-being. This role will focus on building systems to detect unwanted model behaviors and prevent disallowed use of models. You will apply your technical skills to uphold our principles of safety, transparency, and oversight while enforcing our terms of service and acceptable use policies.
Annual compensation range for this role: $320,000 - $485,000 USD.
As set forth in Anthropic’s Equal Employment Opportunity policy, we do not discriminate on the basis of any protected group status under any applicable law.