Remote AI Safety Red Teamer: Stress-Test Frontier Models

Mercor

Paris

Sur place

EUR 83 000 - 100 000

Plein temps

Il y a 10 jours

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Résumé du poste

Mercor is seeking an AI Safety Red Teamer to work remotely on high-stakes AI safety projects. You will design adversarial prompts to stress-test frontier AI models and identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.

You will evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains, and document vulnerabilities while contributing to safety benchmarking and red-teaming reports.

Qualifications

  • Bachelor's degree or higher in a relevant field is required.
  • 5+ years of professional AI safety or red-teaming experience.
  • Strong analytical reasoning and prompt-design abilities.
  • Excellent written communication skills.

Responsabilités

  • Design adversarial prompts to stress-test frontier AI models.
  • Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber, biosecurity, and political content.
  • Document vulnerabilities and support safety benchmarking and red-teaming reports.
  • Collaborate with AI researchers to improve alignment and safety.

Connaissances

Analytical reasoning
Prompt design
Written communication

Formation

Bachelor's degree or higher in CS, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or related discipline

Description du poste

Mercor is seeking an AI Safety Red Teamer to work remotely on high-stakes AI safety projects. You will design adversarial prompts to stress-test frontier AI models and identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.

You will evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains, and document vulnerabilities while contributing to safety benchmarking and red-teaming reports.

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

Frontier AI Safety Red Team Specialist
Frontier AI Safety Red Team Specialist

Mercor • Paris

Sur place
EUR 90 000 - 130 000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Mercor • Paris

Sur place
EUR 90 000 - 130 000
Remote AI Safety Evaluator (Contract)
Remote AI Safety Evaluator (Contract)

Visa Hunt • France

À distance
EUR 71 000 - 83 000
AI Safety Practitioner: Expert Evaluator & Alignment
AI Safety Practitioner: Expert Evaluator & Alignment

Mercor • Paris

Sur place
EUR 70 000 - 110 000
Frontier AI Safety Evaluator
Frontier AI Safety Evaluator

Mercor • Paris

Sur place
EUR 70 000 - 110 000
Senior AI Security Pentester & Red Team Engineer
Senior AI Security Pentester & Red Team Engineer

Mistral-Ai • Paris

Sur place
EUR 90 000 - 140 000
Frontier AI Safety Evaluator & Policy Reviewer
Frontier AI Safety Evaluator & Policy Reviewer

Obsidian • Paris

Sur place
EUR 70 000 - 90 000
AI Safety Specialist - Evaluation Expert
AI Safety Specialist - Evaluation Expert

Obsidian • Paris

Sur place
EUR 70 000 - 90 000
AI Safety Specialist - Evaluation Expert
AI Safety Specialist - Evaluation Expert

Mercor • Paris

Sur place
EUR 70 000 - 110 000
AI Safety Practitioner - Expert Evaluator
AI Safety Practitioner - Expert Evaluator

Mercor • Paris

Sur place
EUR 70 000 - 110 000