Threat Intel Manager, Model Exploitation & Fraud

Anthropic

California (MO)

Hybrid

USD 375,000 - 455,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Anthropic is seeking a Threat Intel Manager to build and run the Model Exploitation & Fraud team within Threat Intelligence. You will set strategy, hire and lead a small team of investigators, and design scalable processes and partnerships.

You will direct complex investigations across model distillation, unauthorized AI R&D usage, and fraud networks, while coordinating with policy, enforcement, and engineering to translate findings into bans and safety-by-design improvements.

Qualifications

  • Led investigative, fraud, or threat intelligence teams with senior ICs
  • Fluent in scaled abuse, fraud patterns, and account abuse
  • Proficient in SQL and Python to review data-heavy cases
  • Experience tracking threat actors across surface, deep, and dark web
  • Familiar with large language models and model exploitation risks
  • Built processes or programs from scratch and measured impact

Responsibilities

  • Own strategy, priorities, and outcomes for the threat mission area
  • Hire, manage, and develop a team of technical investigators
  • Define gaps between management and senior ICs with clear lanes
  • Lead complex investigations independently
  • Direct investigations into model distillation, unauthorized AI usage, access, and fraud networks
  • Redesign triage for high-volume detection and build abuse signals
  • Expand coverage to fraud and scams and develop playbooks
  • Own external engagement with US government partners and peers
  • Anticipate changes from resellers and third-party platforms
  • Collaborate with policy, enforcement, and engineering on mitigations

Skills

Team leadership
Threat intelligence
SQL
Python
Stakeholder management

Job description

About Anthropic

Anthropic's mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.

About the role

We are looking for a threat intel manager to build and run our Model Exploitation & Fraud team within Threat Intelligence. This team detects, investigates, and disrupts the large-scale exploitation of Anthropic's AI systems. Model distillation, unauthorized access, account farming and reseller abuse, and fraud and scam operations.

You will set the strategy for the mission area, hire and lead a small team of technical investigators, and build the systems, processes, and partnerships that let the team scale. The team includes established senior investigators who own our deepest technical casework, tracing distillation networks, reseller and proxy ecosystems, and financially motivated actors across first-party surfaces and third-party platforms; your job is to direct, resource, and amplify that work, not duplicate it. This area carries regular U.S. government engagement and requires deeply understanding external black market ecosystems and how they interact with our systems.

Important context: In this position you may be exposed to explicit content spanning a range of topics, including those of a sexual, violent, or psychologically disturbing nature. This role may require responding to escalations during weekends and holidays.

Key responsibilities
  • Own strategy, priorities, and outcomes for the Model Exploitation & Fraud mission area; define what we detect, investigate, action, and share
  • Hire, manage, and develop a team of technical threat investigators; set the quality bar for casework and intelligence reporting
  • Design clear lanes between this role and the team's senior individual contributors: strategy, people leadership, and program ownership sit with you, while ownership of the deepest technical investigations and tradecraft stays with the senior experts closest to the work
  • Capable of independently leading complex investigations.
  • Direct, prioritize, and resource complex investigations into model distillation, unauthorized AI R&D usage, unauthorized access, coordinated account abuse, and fraud/scam networks, partnering with the senior investigators who lead the deepest technical casework and clearing blockers from their path
  • Drive the redesign of triage for a very high-volume detection pipeline: partner with investigators and engineering to build abuse signals, clustering, and agentic investigation workflows that separate sophisticated actors from noise
  • Expand the team's coverage into fraud and scams, building the detection and investigation playbooks from the ground up
  • Own the external engagement program for the area, including regular intelligence sharing with U.S. government partners and industry peers, ensuring the investigators driving the work are visible in those channels
  • Anticipate how resellers, proxies, and third-party platforms change the abuse surface, and shape coverage accordingly
  • Work with policy, enforcement, and engineering to convert findings into bans, product mitigations, and safety-by-design improvements
  • Define and report the team's metrics; brief Safeguards and company leadership on the threat landscape
Minimum qualifications
  • Have led and managed investigative, fraud, platform integrity, or threat intelligence teams, ideally ones built around senior, deeply specialized individual contributors
  • Have strong domain fluency in scaled abuse — fraud patterns, account abuse, unauthorized access, or platform exploitation economics — sufficient to set priorities, pressure-test findings, and earn the confidence of expert investigators
  • Are proficient enough in SQL and Python to review data-heavy casework, pressure-test conclusions, and provide surge capacity when the team needs it
  • Have experience overseeing investigations that track threat actors across surface, deep, and dark web environments, including reseller and access-broker communities
  • Have working familiarity with large language models and a strong grasp of how models can be distilled, extracted, or exploited at scale
  • Have built processes, detection systems, or programs from scratch and can show what changed because of them
  • Communicate crisply with executives, engineers, and external partners alike
Preferred qualifications
  • Experience at a major technology platform on trust and safety, fraud, or abuse investigations at scale
  • Background in financial crime investigation or fraud analytics
  • Experience working directly with U.S. government stakeholders on threat reporting
  • A track record of partnering with, growing, and retaining senior technical specialists, including defining clear scope between management and senior IC tracks
  • Fluency in Mandarin Chinese and/or Russian with nuanced regional and geopolitical context
  • Active Top Secret security clearance
The annual compensation range for this role is listed below.

For sales roles, the range provided is the role's On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role.

Annual Salary:

$375,000-$455,000 USD

Logistics

Minimum education: Bachelor's degree or an equivalent combination of education, training, and/or experience

Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience

Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position

Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices.

Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.

How we're different

We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills.

The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.

Come work with us!

Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Threat Intel Manager, Model Exploitation & Fraud
Threat Intel Manager, Model Exploitation & Fraud

Anthropic • San Francisco (CA)

Hybrid
USD 375,000 - 455,000
Threat Intel Manager, Influence Operations & Surveillance
Threat Intel Manager, Influence Operations & Surveillance

Anthropic • California (MO)

Hybrid
USD 375,000 - 455,000
Competitive compensation
Equity donation matching
Generous vacation
+3
Threat Intel Manager, Influence Operations & Surveillance
Threat Intel Manager, Influence Operations & Surveillance

Anthropic • San Francisco (CA)

Hybrid
USD 375,000 - 455,000
Competitive compensation
Equity donation matching
Generous vacation
+3
Security Engineer - Threat Intel
Security Engineer - Threat Intel

Anthropic • New York (NY), San Francisco (CA), Washington

Hybrid
USD 320,000 - 405,000
Equity matching
Generous vacation
Parental leave
+1
Technical Cyber Threat Investigator
Technical Cyber Threat Investigator

Anthropic • Friendly (MD)

Hybrid
USD 230,000 - 290,000
Competitive salary
Generous vacation
Parental leave
+2
1d Anthropic Threat Intel Manager, Model Exploitation & Fraud San Francisco, CA Anthropic 1d Th[...]
1d Anthropic Threat Intel Manager, Model Exploitation & Fraud San Francisco, CA Anthropic 1d Th[...]

Applied Methods Ltd • San Francisco (CA)

On-site
USD 375,000 - 455,000
Office in San Francisco
Flexible working hours
Equity donation matching
+3
Technical Cyber Threat Investigator
Technical Cyber Threat Investigator

Anthropic • Washington

Hybrid
USD 230,000 - 290,000
Technical Cyber Threat Investigator
Technical Cyber Threat Investigator

Menlo Ventures • San Francisco (CA)

On-site
USD 230,000 - 290,000
Competitive compensation
Flexible working hours
Generous vacation and parental leave
Threat Intelligence Engineer
Threat Intelligence Engineer

Anthropic • Friendly (MD)

Hybrid
USD 300,000 - 405,000
Competitive compensation
Flexible working hours
Generous vacation and parental leave
Staff+ Software Engineer, Safeguards Data
Staff+ Software Engineer, Safeguards Data

Anthropic • New York (NY)

Hybrid
USD 320,000 - 485,000
Competitive compensation
Equity donation matching
Generous vacation and parental leave
+2