AI Red Teamer (LLM Generalist) (Fully Remote)

Handshake AI

United States

Remote

USD 140,000 - 200,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Handshake AI is seeking a red-team security researcher to stress-test large language models by crafting creative adversarial prompts and multi-turn scenarios.

You will probe guardrails, study jailbreak techniques, and evaluate model responses against harm taxonomies, helping advance safer frontier AI systems. This role requires comfort with exposure to harmful content and clear written communication with research partners.

Qualifications

  • Hands-on experience with multiple LLMs for testing and evaluation.
  • Intuition for crafting adversarial prompts; jailbreak familiarity is a plus.
  • Creative problem-solving and clear written communication skills.

Responsibilities

  • Craft creative prompts and multi-turn scenarios to stress-test AI guardrails.
  • Discover jailbreak, evasion and prompt-injection techniques.
  • Evaluate and score model responses against structured harm taxonomies and severity rubrics.

Skills

Adversarial prompts
LLM familiarity
Safety evaluation
Written communication

Tools

ChatGPT
Claude
Gemini

Job description

  • Stress-test large language models by intentionally trying to break them.
  • Design creative, adversarial prompts to expose vulnerabilities like unsafe content, bias, and hallucinations.
  • Probe models across risk categories including content safety, CBRN, cybersecurity, and more.

Day-to-Day Responsibilities:

  • Craft creative prompts and multi-turn scenarios to stress-test AI guardrails.
  • Discover ways around safety filters using jailbreak, evasion, and prompt injection techniques.
  • Evaluate and score model responses against structured harm taxonomies and severity rubrics.

Desired Capabilities:

  • Strong hands-on experience using multiple LLMs (ChatGPT, Claude, Gemini, etc.).
  • Intuition for crafting adversarial prompts; familiarity with jailbreak techniques is a plus.
  • Creative, adversarial problem-solving skills and clear written communication.

Content Warning:

  • This role involves regular exposure to harmful content, including violence, self-harm, and child safety scenarios.
  • Candidates must engage with this material professionally and sustainably.
  • Support resources are available.
Handshake AI

Handshake AI partners with leading AI research labs to make models safer and more robust. Our red teaming operations help identify vulnerabilities before they reach users, contributing directly to the responsible development of frontier AI systems.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Red Teamer (LLM Generalist) - Remote
AI Red Teamer (LLM Generalist) - Remote

Apply • Northern (KY)

Hybrid
USD 120,000 - 180,000
AI Red Teamer (LLM Generalist)
AI Red Teamer (LLM Generalist)

Apply • Seattle (WA), Northern (KY)

On-site
USD 120,000 - 180,000
Remote AI Red Teamer - LLM Generalist (Stress-Test)
Remote AI Red Teamer - LLM Generalist (Stress-Test)

Apply • Northern (KY)

Hybrid
USD 120,000 - 180,000
LLM Generalist Red Teamer — Stress-Testing & Safeguards
LLM Generalist Red Teamer — Stress-Testing & Safeguards

Apply • Seattle (WA), Northern (KY)

Hybrid
USD 120,000 - 180,000
LLM Adversarial Tester & Safety Architect
LLM Adversarial Tester & Safety Architect

Handshake AI • United States

Remote
USD 140,000 - 200,000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Obsidian • San Francisco (CA)

On-site
USD 150,000 - 230,000
AI Red Teamer: Offensive Security for AI Systems (Remote)
AI Red Teamer: Offensive Security for AI Systems (Remote)

Handshake • United States

Remote
USD 150,000 - 210,000
AI Red Teamer: Offensive Security & Threat Modeling
AI Red Teamer: Offensive Security & Threat Modeling

Handshake • United States

Remote
USD 140,000 - 210,000
Health Insurance
401k match
Parental leave
+1
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Mercor • New York (NY)

On-site
USD 150,000 - 190,000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Obsidian • New York (NY)

On-site
USD 170,000 - 260,000