Remote AI Safety Red Teamer: Adversarial Testing

United States Digital Space LLC

United States

Remote

USD 96,000 - 116,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Mercor is seeking an AI Safety Red Teamer on a remote contract basis. The role focuses on designing adversarial prompts to stress-test frontier AI models, identifying jailbreaks and unsafe outputs, and evaluating robustness across a range of sensitive domains.

Collaboration with AI researchers will drive improvements in safety and alignment. Ideal candidates bring 5+ years in AI Safety or related fields, with strong analytical and communication skills, and a background in CS or related

Qualifications

  • Must have strong analytical reasoning, prompt design, and written communication skills.
  • Experience designing adversarial prompts or evaluating frontier AI systems.
  • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or related field.

Responsibilities

  • Design adversarial prompts to stress-test frontier AI models.
  • Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
  • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
  • Collaborate with AI researchers to improve model alignment, robustness, and safety.

Skills

Analytical reasoning
Prompt design
Written communication
Adversarial prompts
Red team experience

Education

Bachelor's degree or higher in CS/related field

Job description

Mercor is seeking an AI Safety Red Teamer on a remote contract basis. The role focuses on designing adversarial prompts to stress-test frontier AI models, identifying jailbreaks and unsafe outputs, and evaluating robustness across a range of sensitive domains.

Collaboration with AI researchers will drive improvements in safety and alignment. Ideal candidates bring 5+ years in AI Safety or related fields, with strong analytical and communication skills, and a background in CS or related

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote AI Safety Red Team Expert
Remote AI Safety Red Team Expert

Mercor • New York (NY)

On-site
USD 120,000 - 180,000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Neon • United States

Remote
USD 90,000 - 150,000
Wellness resources
Remote work
Remote AI Safety Red Team Engineer
Remote AI Safety Red Team Engineer

Neon • United States

Remote
USD 120,000 - 190,000
Remote AI Red Team Safety Specialist: Adversarial Testing
Remote AI Red Team Safety Specialist: Adversarial Testing

Obsidian • San Francisco (CA)

Remote
USD 90,000 - 130,000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • New York (NY)

On-site
USD 110,000 - 180,000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • San Francisco (CA)

On-site
USD 120,000 - 180,000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • San Francisco (CA)

Remote
USD 120,000 - 180,000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • New York (NY)

On-site
USD 120,000 - 160,000
Remote Adversarial ML Specialist — AI Safety Red Team
Remote Adversarial ML Specialist — AI Safety Red Team

Obsidian • San Francisco (CA)

Remote
USD 120,000 - 180,000
Remote work
AI Red Team Specialist — Adversarial Testing (Remote)
AI Red Team Specialist — Adversarial Testing (Remote)

Mercor • San Francisco (CA)

Remote
USD 130,000 - 170,000