Frontier AI Safety Red Team Pro

Mercor

Zürich

Vor Ort

CHF 120.000 - 180.000

Vollzeit

14 Tage+

Erhalte mehr Antworten von Arbeitgebern

Versende in nur wenigen Minuten einen passgenauen Lebenslauf.

Zusammenfassung

Mercor is seeking experienced AI Safety Red Teamers to stress-test frontier AI systems through adversarial prompts and rigorous evaluation. You will identify jailbreaks, unsafe behaviours, and policy failures, while assessing model robustness across sensitive domains such as misinformation and biosecurity.

Collaboration with researchers will help advance safety benchmarking and alignment. Applicants should have strong analytical and communication skills, a Bachelor's degree or higher in a

Qualifikationen

  • Bachelor's degree or higher in a field related to CS, cybersecurity or safety.
  • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or related field.
  • Strong analytical reasoning, prompt design, and written communication skills.
  • Experience designing adversarial prompts or evaluating frontier AI systems.

Aufgaben

  • Design adversarial prompts to stress-test frontier AI models.
  • Identify jailbreaks, unsafe behaviours, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
  • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
  • Collaborate with AI researchers to improve model alignment, robustness, and safety.

Kenntnisse

AI Safety
Red Teaming
Adversarial Testing
Prompt Design
Cybersecurity
Analytical Reasoning
Written Communication

Ausbildung

Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline

Tools

Jailbreak testing
Prompt engineering
Adversarial evaluation

Jobbeschreibung

Mercor is seeking experienced AI Safety Red Teamers to stress-test frontier AI systems through adversarial prompts and rigorous evaluation. You will identify jailbreaks, unsafe behaviours, and policy failures, while assessing model robustness across sensitive domains such as misinformation and biosecurity.

Collaboration with researchers will help advance safety benchmarking and alignment. Applicants should have strong analytical and communication skills, a Bachelor's degree or higher in a

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Frontier AI Safety Red Teamer
Frontier AI Safety Red Teamer

Mercor • Zürich

Remote
CHF 120.000 - 170.000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Mercor • Zürich

Vor Ort
CHF 120.000 - 180.000
Frontier AI Safety Evaluator & Policy Risk Analyst
Frontier AI Safety Evaluator & Policy Risk Analyst

Mercor • Zürich

Remote
CHF 120.000 - 180.000
Frontier AI Safety Evaluator & Model Alignment Expert
Frontier AI Safety Evaluator & Model Alignment Expert

Mercor • Zürich

Vor Ort
CHF 120.000 - 180.000
Frontier AI Safety Evaluator & Alignment Specialist
Frontier AI Safety Evaluator & Alignment Specialist

Mercor • Zürich

Remote
CHF 110.000 - 150.000
Advanced AI Safety Evaluator: Frontier Model Alignment
Advanced AI Safety Evaluator: Frontier Model Alignment

Mercor • Zürich

Vor Ort
CHF 120.000 - 180.000
Frontier AI Safety Evaluator & Alignment Expert
Frontier AI Safety Evaluator & Alignment Expert

Obsidian • Zürich

Vor Ort
CHF 120.000 - 170.000
AI Safety Practitioner - Expert Evaluator
AI Safety Practitioner - Expert Evaluator

Mercor • Zürich

Vor Ort
CHF 120.000 - 180.000
AI Safety Specialist - Evaluation Expert
AI Safety Specialist - Evaluation Expert

Mercor • Zürich

Vor Ort
CHF 120.000 - 180.000
AI Safety Specialist - Evaluation Expert
AI Safety Specialist - Evaluation Expert

Obsidian • Zürich

Vor Ort
CHF 120.000 - 170.000