AI Safety Red Teamer

AIUC

Deutschland

Remote

EUR 90.000 - 130.000

Vollzeit

14 Tage+
Bewerbungsgenerator

A complete application in a minute — tailored resume and cover letter, ready to send.

Schaffe es an den ATS-Filtern vorbei

Zusammenfassung

AIUC is seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across high-risk topics.

Responsibilities include stress-testing models, identifying jailbreaks and policy failures, and contributing to safety benchmarking while collaborating with researchers to improve alignment and safety.

Qualifikationen

  • Bachelor’s degree or higher in a relevant field.
  • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or related field.
  • Strong analytical reasoning, prompt design, and written communication skills.
  • Experience designing adversarial prompts or evaluating frontier AI systems.

Aufgaben

  • Design adversarial prompts to stress-test frontier AI models.
  • Identify jailbreaks, unsafe behaviours, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
  • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
  • Collaborate with AI researchers to improve model alignment, robustness, and safety.

Kenntnisse

Adversarial prompt design
Analytical reasoning
Written communication
AI Safety/Red Teaming experience

Ausbildung

Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or related discipline

Jobbeschreibung

We are seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high-risk, and ambiguous ("grey-area") topics.

Responsibilities
  • Design adversarial prompts to stress-test frontier AI models.
  • Identify jailbreaks, unsafe behaviours, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
  • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
  • Collaborate with AI researchers to improve model alignment, robustness, and safety.
Required Qualifications
  • Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline.
  • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field.
  • Strong analytical reasoning, prompt design, and written communication skills.
  • Experience designing adversarial prompts or evaluating frontier AI systems.
Preferred Qualifications
  • Experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety.
  • Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies.
  • Expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety.
Why Join?
  • Help secure and strengthen the next generation of frontier AI models.
  • Work on cutting-edge adversarial testing alongside leading AI researchers and safety teams.
  • Influence how AI systems respond to complex, real-world safety challenges.
Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

AI Safety Practitioner - Fully Remote
AI Safety Practitioner - Fully Remote

Mercor • Berlin

Remote
EUR 90.000 - 130.000
AI Safety Specialist - Evaluation Expert
AI Safety Specialist - Evaluation Expert

Obsidian • Berlin

Vor Ort
EUR 90.000 - 140.000
AI Safety Practitioner - Expert Evaluator
AI Safety Practitioner - Expert Evaluator

Mercor • Berlin

Vor Ort
EUR 70.000 - 120.000
AI Safety Specialist - Evaluation Expert
AI Safety Specialist - Evaluation Expert

Mercor • Berlin

Vor Ort
EUR 75.000 - 110.000
AI Safety Expert - Red Teaming
AI Safety Expert - Red Teaming

Mercor • Berlin

Vor Ort
EUR 70.000 - 110.000
AI Safety Specialist - Fully Remote
AI Safety Specialist - Fully Remote

Mercor • Berlin

Remote
EUR 70.000 - 110.000
Member of Technical Staff - AI Security
Member of Technical Staff - AI Security

Exponential Security Labs GmbH • Tübingen

Hybrid
EUR 90.000 - 120.000
Competitive equity
Remote-friendly setup
Conference budget
+1
Freelance AI Red Team Engineer
Freelance AI Red Team Engineer

Mindrift • Berlin

Vor Ort
EUR 54.611 - 77.780
Competitive hourly rates
Flexible working hours
Experience with advanced AI projects
Research Intern - AI Security
Research Intern - AI Security

Exponential Security Labs GmbH • Tübingen

Hybrid
EUR 8.900 - 13.000
Remote-friendly setup
Conference budget
Learning budget
+1
Cybersecurity Expert - Offensive Security
Cybersecurity Expert - Offensive Security

Mercor • Berlin

Vor Ort
EUR 90.000 - 130.000