AI Safety Specialist - Fully Remote | Upto $84/hr

mercor

Italia

Remote

EUR 178,223,000 - 215,096,000

Part time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Mercor is seeking an AI Safety Red Teamer for a remote contract role. The position focuses on stress-testing frontier AI models through adversarial prompts, identifying jailbreaks, unsafe behaviors, and policy failures.

You will evaluate robustness across domains like misinformation, cyber, biosecurity, and political content, document vulnerabilities, and contribute to red-teaming reports. Collaboration with researchers will improve alignment and safety.

Qualifications

  • Bachelor's degree or higher in CS, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline.
  • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field.
  • Strong analytical reasoning, prompt design, and written communication skills.
  • Experience designing adversarial prompts or evaluating frontier AI systems.

Responsibilities

  • Design adversarial prompts to stress-test frontier AI models.
  • Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
  • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
  • Collaborate with AI researchers to improve model alignment, robustness, and safety.

Skills

Adversarial prompts
AI safety
Red team
Prompt design
Written communication

Education

Bachelor's degree in CS or related

Job description

About the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey .

Position: AI Safety Red Teamer | Type: Contract | Compensation: $70–$84/hour | Location: Remote

Role Responsibilities
  • Design adversarial prompts to stress-test frontier AI models .
  • Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
  • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
  • Collaborate with AI researchers to improve model alignment, robustness, and safety.
Qualifications Must-Have
  • Bachelor's degree or higher in Computer Science , Cybersecurity , Journalism , Communications , Psychology , Biology , Chemistry , Public Policy , or a related discipline.
  • 5+ years of professional experience in AI Safety , AI Red Teaming , Trust & Safety , cybersecurity, investigative journalism, life sciences, or a related field.
  • Strong analytical reasoning, prompt design, and written communication skills.
  • Experience designing adversarial prompts or evaluating frontier AI systems .
Preferred
  • Experience with AI Red Teaming , RLHF , SFT , AI Alignment , or Trust & Safety .
  • Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies.
  • Expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

mercor • Italy

Remote
EUR 86,000 - 103,000
AI Safety Practitioner - Fully Remote | Upto $70/hr
AI Safety Practitioner - Fully Remote | Upto $70/hr

mercor • Italy

Remote
EUR 74,000 - 86,000
AI Safety Specialist - Evaluation Expert
AI Safety Specialist - Evaluation Expert

mercor • Italy

Remote
EUR 74,000 - 86,000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Mercor • Lombardia

On-site
EUR 90,000 - 130,000
Remote AI Safety Red Teamer - Adversarial Prompting
Remote AI Safety Red Teamer - Adversarial Prompting

Mercor • Roma

On-site
EUR 84,000 - 100,000
AI Safety Evaluator - Remote Contract
AI Safety Evaluator - Remote Contract

mercor • Italy

Remote
EUR 74,000 - 86,000
AI Safety Evaluator & Model Alignment Specialist
AI Safety Evaluator & Model Alignment Specialist

mercor • Italy

Remote
EUR 74,000 - 86,000
Bilingual Content Specialist - Fully Remote | Upto $44/hr Part-time
Bilingual Content Specialist - Fully Remote | Upto $44/hr Part-time

mercor • Italy

Remote
EUR 49,000 - 54,000
Financial Analyst - Fully Remote | Upto $130/hr
Financial Analyst - Fully Remote | Upto $130/hr

mercor • Italy

Remote
EUR 123,000 - 160,000
Freelance Cybersecurity Analyst - AI Trainer
Freelance Cybersecurity Analyst - AI Trainer

Mindrift • Roma

On-site
EUR 32,861 - 46,802
Competitive pay rates up to $34/hour
Flexible freelance work hours
Experience on advanced AI projects