AI Safety Specialist - Remote

Mercor

Lazio

Remoto

EUR 84.000 - 101.000

Part-time

14 giorni+
Generatore di candidature

Trasforma questo ruolo in un colloquio — un curriculum e una lettera di presentazione personalizzati in base a ciò che cerca questo datore di lavoro.

Supera i filtri ATS

Descrizione del lavoro

Mercor is seeking an AI Safety Red Teamer for a remote contract role. You will design adversarial prompts to stress-test frontier AI models and identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.

You will collaborate with AI researchers to improve alignment, robustness, and safety, documenting vulnerabilities and contributing to safety benchmarking. A strong background in AI safety or red-teaming and a Bachelor's degree are required.

Competenze

  • Bachelor's degree or higher in CS, cybersecurity, journalism, communications, psychology, biology, chemistry, public policy or related discipline.
  • 5+ years in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences or related field.
  • Strong analytical reasoning, prompt design, and written communication.
  • Experience designing adversarial prompts or evaluating frontier AI systems.

Mansioni

  • Design adversarial prompts to stress-test frontier AI models.
  • Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber security, biosecurity, fraud, political content, and other sensitive domains.
  • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
  • Collaborate with AI researchers to improve model alignment, robustness, and safety.

Conoscenze

Analytical skills
Prompt design
Written communication
Adversarial prompts
AI safety

Formazione

Bachelor's degree or higher

Descrizione del lavoro

About the job

Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D'Angelo, Larry Summers, and Jack Dorsey.

Position: AI Safety Red Teamer
Type: Contract
Compensation: $70–$84/hour
Location: Remote

Role Responsibilities
  • Design adversarial prompts to stress-test frontier AI models.
  • Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
  • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
  • Collaborate with AI researchers to improve model alignment, robustness, and safety.
Qualifications
Must-Have
  • Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline.
  • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field.
  • Strong analytical reasoning, prompt design, and written communication skills.
  • Experience designing adversarial prompts or evaluating frontier AI systems.
Preferred
  • Experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety.
  • Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies.
  • Expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety.
Ottieni la revisione del curriculum gratis e riservata.

o trascina qui il file.

Similar jobs

Offerte di lavoro simili che vale la pena confrontare

Remote AI Safety Red Teamer (Contract)
Remote AI Safety Red Teamer (Contract)

Mercor • Lazio

Remoto
EUR 84.000 - 101.000
Frontier AI Safety Red Teamer - Adversarial Testing Lead
Frontier AI Safety Red Teamer - Adversarial Testing Lead

Mercor • Milano

In loco
EUR 90.000 - 130.000
Remote AI Safety Red Teamer - Adversarial Prompting
Remote AI Safety Red Teamer - Adversarial Prompting

Mercor • Roma

In loco
EUR 84.000 - 100.000
Freelance Cybersecurity Analyst - AI Trainer
Freelance Cybersecurity Analyst - AI Trainer

Mindrift • Roma

In loco
EUR 32.861 - 46.802
Competitive pay rates up to $34/hour
Flexible freelance work hours
Experience on advanced AI projects
Frontier AI Safety Evaluator
Frontier AI Safety Evaluator

Mercor • Roma

Remoto
EUR 60.000 - 90.000
AI Safety Researcher
AI Safety Researcher

CTI Clinical Trial and Consulting Services • Roma

In loco
EUR 60.000 - 90.000
Remote AI Safety Tester
Remote AI Safety Tester

TSMG Holding • Roma

Remoto
EUR 26.000 - 39.000
Senior AI Full-Stack Engineer (Angular&Python)
Senior AI Full-Stack Engineer (Angular&Python)

NTT DATA Europe & Latam • Emilia-Romagna

In loco
EUR 85.000 - 120.000
AI Agent Evaluation Analyst (Freelance)
AI Agent Evaluation Analyst (Freelance)

Mindrift • Milano

In loco
EUR 23.623 - 35.435
Flexible, remote work
Competitive pay up to $30/hour
Gain experience in advanced AI projects
Frontier AI Safety Evaluator: Expert Feedback
Frontier AI Safety Evaluator: Expert Feedback

Mercor • Roma

In loco
EUR 65.000 - 95.000