Remote AI Safety Red Teamer for Frontier Models

Mercor

Helsinki

Remote

EUR 90,000 - 120,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Mercor in Helsinki seeks experienced AI Safety Red Teamers to stress-test frontier AI systems. You will design challenging prompts, uncover model weaknesses, and assess behavior across high-risk topics.

Responsibilities include identifying jailbreaks, unsafe behaviours, hallucinations, and policy failures, and documenting vulnerabilities for safety benchmarking. You will collaborate with AI researchers to improve alignment and robustness.

Qualifications

  • Bachelor's degree or higher in a related field.
  • 5+ years of professional AI safety or red teaming experience.
  • Strong analytical reasoning and prompt design skills.
  • Experience designing adversarial prompts or evaluating frontier AI systems.

Responsibilities

  • Design adversarial prompts to stress-test frontier AI models.
  • Identify jailbreaks, unsafe behaviours, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
  • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
  • Collaborate with AI researchers to improve model alignment, robustness, and safety.

Education

Bachelor's degree or higher in CS, cybersecurity, journalism, communications, psychology, biology, chemistry, public policy, or related discipline
5+ years of AI safety/red teaming experience
Prompt design and analytical reasoning
Experience designing adversarial prompts
Written communication skills

Job description

Mercor in Helsinki seeks experienced AI Safety Red Teamers to stress-test frontier AI systems. You will design challenging prompts, uncover model weaknesses, and assess behavior across high-risk topics.

Responsibilities include identifying jailbreaks, unsafe behaviours, hallucinations, and policy failures, and documenting vulnerabilities for safety benchmarking. You will collaborate with AI researchers to improve alignment and robustness.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Frontier AI Safety Red Teamer
Frontier AI Safety Red Teamer

Mercor • Helsinki

Remote
EUR 90,000 - 140,000
Frontier AI Safety Red Team Specialist
Frontier AI Safety Red Team Specialist

Mercor • Helsinki

On-site
EUR 100,000 - 180,000
Frontier AI Safety Red Teamer
Frontier AI Safety Red Teamer

Obsidian • Helsinki

Remote
EUR 90,000 - 130,000
Remote AI Safety Red Team Engineer (English/Finnish)
Remote AI Safety Red Team Engineer (English/Finnish)

Mercor • Helsinki

On-site
EUR 90,000 - 130,000
Remote AI Safety Red Team Specialist (English & Finnish)
Remote AI Safety Red Team Specialist (English & Finnish)

Obsidian • Helsinki

Remote
EUR 77,000 - 112,000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Mercor • Helsinki

On-site
EUR 100,000 - 180,000
Remote AI Safety Red Team Expert
Remote AI Safety Red Team Expert

Mercor • Helsinki

On-site
EUR 85,000 - 120,000
Remote-friendly operations
Wellness resources and guidelines
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • Helsinki

Remote
EUR 77,000 - 111,000
Remote AI Adversarial Red Team Specialist (EN/FI)
Remote AI Adversarial Red Team Specialist (EN/FI)

Obsidian • Helsinki

Remote
EUR 78,000 - 113,000
AI Safety Red Teamer - Remote Contract
AI Safety Red Teamer - Remote Contract

Mercor • Helsinki

On-site
EUR 83,000 - 100,000