AI Safety Specialist - Fully Remote | Upto $84/hr

Visa Hunt

Denmark

Remote

DKK 621,000 - 745,000

Part time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Mercor is seeking an AI Safety Red Teamer to stress-test frontier AI models, identify jailbreaks and policy failures, and collaborate with AI researchers on alignment and safety.

This contract role is remote and pays $70–$84/hour, requiring a Bachelor's degree in a related field and 5+ years in AI safety or related disciplines.

Qualifications

  • Bachelor's degree or higher in a related field.
  • 5+ years of professional experience in AI Safety, AI Red Teeing? cuff? Trust & Safety or related field.
  • Strong analytical reasoning, prompt design, and written communication skills.
  • Experience designing adversarial prompts or evaluating frontier AI systems.

Responsibilities

  • Design adversarial prompts to stress-test frontier AI models.
  • Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
  • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
  • Collaborate with AI researchers to improve model alignment, robustness, and safety.

Skills

Adversarial prompts design
AI safety analysis
Written communication

Education

Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or related discipline

Job description

About the job

Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey.


Position: AI Safety Red Teamer
Type:Contract
Compensation:$70–$84/hour
Location:Remote


Role Responsibilities


  • Design adversarial prompts to stress-test frontier AI models.

  • Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.

  • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.

  • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.

  • Collaborate with AI researchers to improve model alignment, robustness, and safety.


Qualifications

Must-Have



  • Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline.

  • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field.

  • Strong analytical reasoning, prompt design, and written communication skills.

  • Experience designing adversarial prompts or evaluating frontier AI systems.


Preferred


  • Experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety.

  • Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies.

  • Expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Mercor • København

On-site
DKK 900,000 - 1,500,000
AI Safety Specialist - Fully Remote
AI Safety Specialist - Fully Remote

Mercor • København

Remote
DKK 578,000 - 963,000
AI Safety Red Team Lead — Frontier Models
AI Safety Red Team Lead — Frontier Models

Mercor • København

Remote
DKK 850,000 - 1,200,000
Remote AI Red Team Specialist - Adversarial Safety
Remote AI Red Team Specialist - Adversarial Safety

Obsidian • København

On-site
DKK 578,000 - 770,000
AI Safety Specialist - Evaluation Expert
AI Safety Specialist - Evaluation Expert

Mercor • København

On-site
DKK 750,000 - 1,000,000
AI Safety Specialist - Evaluation Expert
AI Safety Specialist - Evaluation Expert

Obsidian • København

On-site
DKK 900,000 - 1,300,000
Frontier AI Safety Red Team Expert
Frontier AI Safety Red Team Expert

Mercor • København

On-site
DKK 900,000 - 1,500,000
AI Safety Practitioner - Expert Evaluator
AI Safety Practitioner - Expert Evaluator

Mercor • København

On-site
DKK 900,000 - 1,200,000
AI Safety Evaluation Specialist – Remote
AI Safety Evaluation Specialist – Remote

Visa Hunt • Denmark

On-site
DKK 533,000 - 621,000
Remote AI Safety Red Team Specialist (English & Danish)
Remote AI Safety Red Team Specialist (English & Danish)

Visa Hunt • Denmark

On-site
DKK 426,000 - 550,000