Remote AI Red Teamer - LLM Generalist (Stress-Test)

Apply

Northern (KY)

Hybrid

USD 120,000 - 180,000

Part time

12 days ago
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Handshake AI seeks an AI Red Teamer to stress-test large language models by crafting adversarial prompts that reveal vulnerabilities in safety guardrails, bias, and jailbreaking weaknesses. This role covers multiple risk areas including content safety, cybersecurity, and high-risk domains, across text, image, and agentic capabilities.

You will work with engineers and researchers to document experiments, score outputs against harm taxonomies, and help strengthen defenses while operating with

Qualifications

  • Stress-test AI models by crafting creative adversarial prompts.
  • Evaluate model outputs against harm taxonomies and severity rubrics.
  • Document experiments with methods and findings.
  • Collaborate with engineers and researchers to strengthen defenses.

Responsibilities

  • Craft prompts and multi-turn scenarios to stress-test AI guardrails across diverse risk categories.
  • Discover jailbreak, evasion, and prompt-injection techniques.
  • Explore edge cases to provoke disallowed or harmful outputs.
  • Evaluate and score model responses using structured harm taxonomies.

Skills

LLM experience
Adversarial prompts
Creative problem-solving
Written communication
Ethical judgment
Self-directed
Curiosity
Python basics
LLM APIs
Data annotation
Trust and safety
Subject matter expertise

Tools

Python

Job description

Handshake AI seeks an AI Red Teamer to stress-test large language models by crafting adversarial prompts that reveal vulnerabilities in safety guardrails, bias, and jailbreaking weaknesses. This role covers multiple risk areas including content safety, cybersecurity, and high-risk domains, across text, image, and agentic capabilities.

You will work with engineers and researchers to document experiments, score outputs against harm taxonomies, and help strengthen defenses while operating with

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

LLM Generalist Red Teamer — Stress-Testing & Safeguards
LLM Generalist Red Teamer — Stress-Testing & Safeguards

Apply • Seattle (WA), Northern (KY)

Hybrid
USD 120,000 - 180,000
AI Red Teamer: LLM Safety & Stress-Testing Generalist
AI Red Teamer: LLM Safety & Stress-Testing Generalist

Handshake • Seattle (WA)

On-site
USD 83,000 - 138,000
AI Red Teamer (LLM Generalist) - Remote
AI Red Teamer (LLM Generalist) - Remote

Apply • Northern (KY)

Hybrid
USD 120,000 - 180,000
AI Red Teamer (LLM Generalist)
AI Red Teamer (LLM Generalist)

Handshake • Seattle (WA)

On-site
USD 83,000 - 138,000
AI Red Teamer (LLM Generalist)
AI Red Teamer (LLM Generalist)

Apply • Seattle (WA), Northern (KY)

On-site
USD 120,000 - 180,000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • New York (NY)

On-site
USD 90,000 - 140,000
Remote AI Red Team Engineer
Remote AI Red Team Engineer

Mercor • New York (NY)

Remote
USD 90,000 - 160,000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • San Francisco (CA)

Remote
USD 120,000 - 180,000
Remote AI Safety Red Team Expert (Adversarial ML)
Remote AI Safety Red Team Expert (Adversarial ML)

Mercor • New York (NY)

Remote
USD 110,000 - 170,000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • New York (NY)

On-site
USD 120,000 - 160,000