LLM Generalist Red Teamer — Stress-Testing & Safeguards

Apply

Seattle, Northern (WA, KY)

Hybrid

USD 120,000 - 180,000

Full time

8 days ago
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Handshake AI is seeking an AI Red Teamer to stress-test large language models by crafting adversarial prompts and evaluating guardrails. You will work across content safety, cybersecurity, and high-risk domains, collaborating with researchers to strengthen defenses.

The role involves regular exposure to potentially disturbing content and requires strong documentation, written communication, and ethical judgment. Seattle-based, contract, 40 hours/week.

Qualifications

  • Experience stress-testing multiple LLMs and evaluating guardrails.
  • Ability to design adversarial prompts and jailbreak tactics.
  • Ability to document experiments clearly and consistently.

Responsibilities

  • Craft adversarial prompts to stress-test AI guardrails.
  • Discover jailbreak and prompt-injection techniques.
  • Probe edge cases to provoke disallowed outputs.
  • Evaluate outputs using harm taxonomies and rubrics.
  • Document experiments, including methods and results.
  • Review prompts from teammates.
  • Contribute to harm taxonomy development and calibration.
  • Collaborate with engineers, data scientists, researchers.
  • Work with potentially disturbing content; stay professional.
  • Stay current on jailbreaks and model behaviors.

Skills

LLMs experience
Adversarial prompts
Content safety
Ethical judgment
Written communication
Documentation
Collaboration
Python basics

Tools

LLM APIs
Evaluation tooling
Python scripting

Job description

Handshake AI is seeking an AI Red Teamer to stress-test large language models by crafting adversarial prompts and evaluating guardrails. You will work across content safety, cybersecurity, and high-risk domains, collaborating with researchers to strengthen defenses.

The role involves regular exposure to potentially disturbing content and requires strong documentation, written communication, and ethical judgment. Seattle-based, contract, 40 hours/week.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Red Teamer: LLM Safety & Stress-Testing Generalist
AI Red Teamer: LLM Safety & Stress-Testing Generalist

Handshake • Seattle (WA)

On-site
USD 83,000 - 138,000
AI Red Teamer (LLM Generalist)
AI Red Teamer (LLM Generalist)

Apply • Seattle (WA), Northern (KY)

Hybrid
USD 120,000 - 180,000
AI Red Teamer (LLM Generalist)
AI Red Teamer (LLM Generalist)

Handshake • Seattle (WA)

On-site
USD 83,000 - 138,000
LLM Red-Team Specialist for Adversarial Evaluation
LLM Red-Team Specialist for Adversarial Evaluation

Mercor • New York (NY)

On-site
USD 90,000 - 150,000
Remote work within the United States
Remote AI Safety Red Teamer (Contract)
Remote AI Safety Red Teamer (Contract)

United States Digital Space LLC • United States

Remote
USD 96,000 - 116,000
AI Red Team Analyst (LLM Safety & Adversarial Testing) | $26/hr Remote
AI Red Team Analyst (LLM Safety & Adversarial Testing) | $26/hr Remote

Crossing Hurdles • United States

On-site
MXN 639,828
AI Red Team Engineer — LLM Security & Adversarial Testing
AI Red Team Engineer — LLM Security & Adversarial Testing

Arcitix Security • United States

On-site
USD 100,000 - 140,000
AI Red Teamer: Offensive Security for AI Systems (Remote)
AI Red Teamer: Offensive Security for AI Systems (Remote)

Handshake • United States

Remote
USD 150,000 - 210,000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • New York (NY)

On-site
USD 90,000 - 140,000
Adversarial AI Testing Specialist (LLM Red Teaming) | $28.74/hr Remote
Adversarial AI Testing Specialist (LLM Red Teaming) | $28.74/hr Remote

Crossing Hurdles • United States

Remote