An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Anthropic is seeking software engineers to build safety and oversight mechanisms for AI systems. You will monitor models, prevent misuse, and ensure user wellbeing, focusing on scalable, distributed infrastructure and sandboxed runtimes.
Ideal candidates have strong Python and distributed systems experience, multi-cloud exposure, and a background in security or compliance. Rolling reviews; compensation listed below.
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.
We are looking for software engineers to help build safety and oversight mechanisms for our AI systems. As a software engineer on the Safeguards team, you will work to monitor models, prevent misuse, and ensure user well-being. This role will focus on building distributed and scalable systems to detect unwanted model behaviors and prevent disallowed use of models. You will apply your technical skills to uphold our principles of safety, transparency, and oversight while enforcing our terms of service and acceptable use policies.
Deadline to apply: None. Applications will be reviewed on a rolling basis.
The annual compensation range for this role is listed below.
For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role.
Annual Salary: $320,000 — $485,000 USD
Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience
Required field of study:<