AI Red Teamer: Offensive Security for Safe AI

Handshake

Seattle (WA)

On-site

USD 120,000 - 210,000

Full time

9 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Handshake is seeking a Cybersecurity Red Teamer to evaluate whether AI models can be manipulated into generating malware, exploits, or attacker guidance. You will craft adversarial prompts and multi-turn interactions to simulate threat actor use of LLMs, then assess output risk and exploitability with deep cybersecurity expertise.

Day-to-day work includes designing prompts, evaluating model outputs for functional robustness, testing across multiple offensive categories, and collaborating with

Qualifications

  • Experience in offensive security, penetration testing, red teaming, or incident response.
  • Ability to read, write, and evaluate code across common languages used in offensive tooling.
  • Understanding of MITRE ATT&CK, OWASP, and related frameworks.
  • Ability to assess functional correctness and real-world exploitability of model outputs.
  • Strong written and verbal communication for technical risk framing.

Responsibilities

  • Design adversarial prompts to test AI models across cyber kill chains (reconnaissance to exfiltration).
  • Evaluate model-generated code and outputs for functionality and exploitability.
  • Test across malware, vulnerability exploitation, social engineering, credential harvesting, and exfiltration techniques.
  • Probe dual-use boundaries and assess safe-guard effectiveness.
  • Simulate attacker personas across skill levels to gauge risk escalation.
  • Document findings with clear reasoning on what helps or harms attackers.

Skills

Offensive security
Penetration testing
Red teaming
Threat modeling
Adversarial thinking
Written communication

Education

Security certifications (OSCP/OSCE/GPEN)

Tools

Python
PowerShell
Bash
C/C++
LLM tooling

Job description

Handshake is seeking a Cybersecurity Red Teamer to evaluate whether AI models can be manipulated into generating malware, exploits, or attacker guidance. You will craft adversarial prompts and multi-turn interactions to simulate threat actor use of LLMs, then assess output risk and exploitability with deep cybersecurity expertise.

Day-to-day work includes designing prompts, evaluating model outputs for functional robustness, testing across multiple offensive categories, and collaborating with

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Security Red Teamer: Offensive AI Risk & Defense
AI Security Red Teamer: Offensive AI Risk & Defense

Dorado • Seattle (WA)

On-site
USD 160,000 - 260,000
AI Red Team Engineer
AI Red Team Engineer

Confidential • San Francisco (CA)

On-site
USD 120,000 - 160,000
AI Red Teamer, Cybersecurity
AI Red Teamer, Cybersecurity

Handshake • Seattle (WA)

On-site
USD 120,000 - 210,000
AI Red Team Lead — Offensive Security & TTPs
AI Red Team Lead — Offensive Security & TTPs

Jobtailor • Pennsylvania

On-site
USD 140,000 - 190,000
AI Red Teamer: LLM Generalist for Safety Testing
AI Red Teamer: LLM Generalist for Safety Testing

Handshake • Seattle (WA)

On-site
USD 90,000 - 130,000
AI Red Teamer, Cybersecurity
AI Red Teamer, Cybersecurity

Dorado • Seattle (WA)

On-site
USD 160,000 - 260,000
AI Red Team Tester — Adversarial Testing & Security Research
AI Red Team Tester — Adversarial Testing & Security Research

Orion Innovation • United States

Remote
USD 120,000 - 180,000
AI Red Teaming Engineer
AI Red Teaming Engineer

DeWinter Group • Campbell (CA)

Remote
Remote AI Red Team Engineer
Remote AI Red Team Engineer

Mercor • New York (NY)

Remote
USD 90,000 - 160,000
AI Safety Red Team Engineer
AI Safety Red Team Engineer

Anthropic • San Francisco (CA)

Hybrid
USD 320,000 - 405,000