Remote AI Safety Red Teamer — Stress-Test Frontier Models

mercor

United States

Remote

USD 96,000 - 116,000

Part time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Mercor is seeking an AI Safety Red Teamer for a remote, contract role with hourly compensation of $70–$84. You will design adversarial prompts, identify jailbreaks and unsafe behaviors, and test model robustness across sensitive domains.

You will collaborate with AI researchers to improve alignment and safety. The ideal candidate has 5+ years in AI safety or red teaming, a Bachelor's degree or higher in CS or a related field, and strong analytical, prompt design, and written communication

Qualifications

  • Must have a Bachelor’s degree or higher in a related field.
  • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences or related.
  • Strong analytical reasoning, prompt design, and written communication skills.
  • Experience designing adversarial prompts or evaluating frontier AI systems.

Responsibilities

  • Design adversarial prompts to stress-test frontier AI models.
  • Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
  • Document vulnerabilities and contribute to safety benchmarking and red‑teaming reports.
  • Collaborate with AI researchers to improve model alignment, robustness, and safety.

Skills

Prompt design
Analytical reasoning
Written communication
AI safety
Red teaming

Education

Bachelor's degree or higher

Tools

Jailbreak testing
Prompt engineering
Adversarial evaluation

Job description

Mercor is seeking an AI Safety Red Teamer for a remote, contract role with hourly compensation of $70–$84. You will design adversarial prompts, identify jailbreaks and unsafe behaviors, and test model robustness across sensitive domains.

You will collaborate with AI researchers to improve alignment and safety. The ideal candidate has 5+ years in AI safety or red teaming, a Bachelor's degree or higher in CS or a related field, and strong analytical, prompt design, and written communication

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote AI Safety Red Teamer (Contract)
Remote AI Safety Red Teamer (Contract)

United States Digital Space LLC • United States

Remote
USD 96,000 - 116,000
Remote AI Safety Red Teamer for Frontier Models
Remote AI Safety Red Teamer for Frontier Models

Mercor • San Francisco (CA)

Remote
USD 170,000 - 260,000
Remote AI Safety Red Team Expert
Remote AI Safety Red Team Expert

Mercor • New York (NY)

On-site
USD 120,000 - 180,000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • New York (NY)

On-site
USD 120,000 - 160,000
Remote AI Safety & Red Team Specialist
Remote AI Safety & Red Team Specialist

Mercor • San Francisco (CA)

Remote
USD 120,000 - 180,000
Remote AI Safety Red Team Expert (Adversarial ML)
Remote AI Safety Red Team Expert (Adversarial ML)

Mercor • New York (NY)

Remote
USD 110,000 - 170,000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • New York (NY)

On-site
USD 90,000 - 140,000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • San Francisco (CA)

On-site
USD 120,000 - 180,000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • San Francisco (CA)

Remote
USD 120,000 - 180,000
Remote AI Safety Red Team Specialist (EN/SE)
Remote AI Safety Red Team Specialist (EN/SE)

Mercor • New York (NY)

On-site
USD 120,000 - 180,000