AI Safety Red Teamer Expert

mercor

Italia

À distance

EUR 86 000 - 103 000

Plein temps

14 jours+
Générateur de candidature

Obtenez une réponse de cet employeur — un CV et une lettre de motivation adaptés exactement à ce qu’il recherche.

Passez les filtres ATS

Résumé du poste

Mercor is seeking an AI Safety Red Teamer for a remote contract engagement. You will design adversarial prompts, identify jailbreaks and unsafe model behaviors, and evaluate robustness across sensitive domains.

You will document vulnerabilities and contribute to red-teaming reports, collaborating with AI researchers to improve alignment and safety. Required are a 5+ year track record in AI safety or related fields, and strong skills in analytical reasoning, prompt design, and written

Qualifications

  • Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or related discipline.
  • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or related field.
  • Strong analytical reasoning, prompt design, and written communication skills.
  • Experience designing adversarial prompts or evaluating frontier AI systems.

Responsabilités

  • Design adversarial prompts to stress-test frontier AI models.
  • Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
  • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
  • Collaborate with AI researchers to improve model alignment, robustness, and safety.

Connaissances

Analytical reasoning
Prompt design
Written communication
Adversarial testing
AI safety

Formation

Bachelor's degree or higher in CS / Cybersecurity / related field

Description du poste

About the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey .

Position: AI Safety Red Teamer | Type: Contract | Compensation: $70–$84/hour | Location: Remote

Role Responsibilities
  • Design adversarial prompts to stress-test frontier AI models .
  • Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
  • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
  • Collaborate with AI researchers to improve model alignment, robustness, and safety.
Qualifications Must-Have
  • Bachelor's degree or higher in Computer Science , Cybersecurity , Journalism , Communications , Psychology , Biology , Chemistry , Public Policy , or a related discipline.
  • 5+ years of professional experience in AI Safety , AI Red Teaming , Trust & Safety , cybersecurity, investigative journalism, life sciences, or a related field.
  • Strong analytical reasoning, prompt design, and written communication skills.
  • Experience designing adversarial prompts or evaluating frontier AI systems .
Preferred
  • Experience with AI Red Teaming , RLHF , SFT , AI Alignment , or Trust & Safety .
  • Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies.
  • Expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety.
Obtenez votre examen gratuit et confidentiel de votre CV.

ou faites glisser et déposez votre fichier ici.

Similar jobs

Postes similaires à comparer

AI Safety Specialist - Fully Remote | Upto $84/hr
AI Safety Specialist - Fully Remote | Upto $84/hr

mercor • Italie

À distance
EUR 178 223 000 - 215 096 000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Mercor • Lombardia

Sur place
EUR 90 000 - 130 000
AI Safety Specialist - Evaluation Expert
AI Safety Specialist - Evaluation Expert

mercor • Italie

À distance
EUR 74 000 - 86 000
Remote AI Safety Red Teamer — Stress-Test Frontier Models
Remote AI Safety Red Teamer — Stress-Test Frontier Models

mercor • Italie

À distance
EUR 178 223 000 - 215 096 000
AI Safety Practitioner - Fully Remote | Upto $70/hr
AI Safety Practitioner - Fully Remote | Upto $70/hr

mercor • Italie

À distance
EUR 74 000 - 86 000
Remote AI Safety Red Teamer - Adversarial Prompting
Remote AI Safety Red Teamer - Adversarial Prompting

Mercor • Roma

Sur place
EUR 84 000 - 100 000
AI Safety Evaluator - Remote Contract
AI Safety Evaluator - Remote Contract

mercor • Italie

À distance
EUR 74 000 - 86 000
Remote AI Safety Evaluator (Contract) — Expert Alignment & Risk
Remote AI Safety Evaluator (Contract) — Expert Alignment & Risk

mercor • Italie

À distance
EUR 74 000 - 86 000
AI Safety Evaluator & Model Alignment Specialist
AI Safety Evaluator & Model Alignment Specialist

mercor • Italie

À distance
EUR 74 000 - 86 000
Frontier AI Safety Evaluator
Frontier AI Safety Evaluator

Mercor • Roma

À distance
EUR 60 000 - 90 000