Cyber Safeguards Enforcement Lead

AISafety

Northern, New York (KY, NY)

Hybrid

USD 285,000 - 330,000

Full time

4 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Anthropic is seeking an Enforcement Lead to manage enforcement actions across products, focusing on detecting and mitigating misuse of AI systems in cyber operations. You will lead a team of Cyber Enforcement Analysts and contractors, and shape enforcement frameworks for flagged activity.

Safety and policy enforcement are core to our mission, with weekend/holiday escalations possible. You’ll collaborate across engineering, policy, and legal teams to uphold safe, honest usage of our AI systems.

Qualifications

  • Experience as a people manager.
  • Cybersecurity knowledge including offensive techniques, malware analysis, or vulnerability research.
  • Experience performing content review, abuse investigations, or policy enforcement at volume.
  • Proficiency in SQL and/or Python for data analysis and threat detection.

Responsibilities

  • Manage a team of Cyber Enforcement Analysts and contractors, overseeing the vision of Cyber Enforcement strategy.
  • Create strategies to detect and mitigate potential misuse of AI systems to facilitate cyberattacks, malware creation, exploitation tooling, and related harmful cyber operations.
  • Collaborate with stakeholders regarding novel, ambiguous, or high-severity cases.
  • Collaborate with the Safeguards Policy Design Team on policy gaps surfaced through real enforcement scenarios.
  • Partner with Engineering and Data Science teams to ensure tooling and measurement support enforcement operations.
  • Keep up to date with emerging AI policy enforcement best practices, threat actor tactics, and the evolving cyber threat landscape.
  • Experience as a people manager.
  • Experience in cybersecurity, including knowledge of offensive techniques, exploit development, malware analysis, or vulnerability research.
  • Experience performing content review, abuse investigations, or policy enforcement at volume.
  • Proficiency in SQL and/or Python for data analysis and threat detection.

Skills

SQL
Python

Education

Bachelor’s degree

Job description

Anthropic is seeking an Enforcement Lead to manage enforcement actions across products, focusing on detecting and mitigating misuse of AI systems in cyber operations. You will lead a team of Cyber Enforcement Analysts and contractors, and shape enforcement frameworks for flagged activity.

Safety and policy enforcement are core to our mission, with weekend/holiday escalations possible. You’ll collaborate across engineering, policy, and legal teams to uphold safe, honest usage of our AI systems.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Cyber Safeguards Enforcement Lead
Cyber Safeguards Enforcement Lead

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 280,000
Head of Cyber Safeguards & Enforcement
Head of Cyber Safeguards & Enforcement

Anthropic • Washington

Hybrid
USD 285,000 - 330,000
Safeguards Enforcement Lead, Cyber Harms — Flexible Hours
Safeguards Enforcement Lead, Cyber Harms — Flexible Hours

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000
Competitive compensation
Equity donation matching
Generous vacation
+3
Remote AI Safeguards & Cyber Enforcement Analyst
Remote AI Safeguards & Cyber Enforcement Analyst

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000
Equity donations
Flexible hours
Vacation policy
+2
AI Safeguards Enforcement Specialist
AI Safeguards Enforcement Specialist

Anthropic • Washington

Hybrid
USD 285,000 - 330,000
AI Safeguards Enforcement Lead
AI Safeguards Enforcement Lead

United States Digital Space LLC • United States

Hybrid
USD 285,000 - 330,000
Cyber Harm Safeguards Analyst — Remote
Cyber Harm Safeguards Analyst — Remote

Appliedmethods • San Francisco (CA), Northern (KY)

Hybrid
USD 285,000 - 330,000
AI Safeguards Enforcement Analyst
AI Safeguards Enforcement Analyst

Anthropic • New York (NY)

Hybrid
USD 285,000 - 330,000
Safeguards Enforcement Lead, Cyber Harms
Safeguards Enforcement Lead, Cyber Harms

EngineersOfAI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 280,000
AI Integrity & Safeguards Enforcement Analyst
AI Integrity & Safeguards Enforcement Analyst

Anthropic • San Francisco (CA)

Hybrid
USD 285,000 - 330,000