Remote AI Safety Red Teamer — Stress-Test Frontier Models

mercor

Italia

Teletrabalho

EUR 178 223 000 - 215 096 000

Tempo parcial

14 dias+
Gerador de candidaturas

Transforma esta função numa entrevista — um currículo e uma carta de apresentação criados à volta do que este empregador procura.

Ultrapassa os filtros ATS

Resumo da oferta

Mercor is seeking an AI Safety Red Teamer for a remote contract role. The position focuses on stress-testing frontier AI models through adversarial prompts, identifying jailbreaks, unsafe behaviors, and policy failures.

You will evaluate robustness across domains like misinformation, cyber, biosecurity, and political content, document vulnerabilities, and contribute to red-teaming reports. Collaboration with researchers will improve alignment and safety.

Qualificações

  • Bachelor's degree or higher in CS, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline.
  • 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field.
  • Strong analytical reasoning, prompt design, and written communication skills.
  • Experience designing adversarial prompts or evaluating frontier AI systems.

Responsabilidades

  • Design adversarial prompts to stress-test frontier AI models.
  • Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.
  • Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
  • Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
  • Collaborate with AI researchers to improve model alignment, robustness, and safety.

Conhecimentos

Adversarial prompts
AI safety
Red team
Prompt design
Written communication

Formação académica

Bachelor's degree in CS or related

Descrição da oferta de emprego

Mercor is seeking an AI Safety Red Teamer for a remote contract role. The position focuses on stress-testing frontier AI models through adversarial prompts, identifying jailbreaks, unsafe behaviors, and policy failures.

You will evaluate robustness across domains like misinformation, cyber, biosecurity, and political content, document vulnerabilities, and contribute to red-teaming reports. Collaboration with researchers will improve alignment and safety.

Obtém a tua avaliação gratuita e confidencial do currículo.

ou arrasta e larga o ficheiro aqui.

Similar jobs

Ofertas semelhantes que vale a pena comparar

Remote AI Safety Red Teamer - Adversarial Prompting
Remote AI Safety Red Teamer - Adversarial Prompting

Mercor • Roma

Presencial
EUR 84 000 - 100 000
Remote AI Safety Red Team Expert (Contract)
Remote AI Safety Red Team Expert (Contract)

mercor • Itália

Teletrabalho
EUR 86 000 - 103 000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

mercor • Itália

Teletrabalho
EUR 86 000 - 103 000
AI Safety Specialist - Fully Remote | Upto $84/hr
AI Safety Specialist - Fully Remote | Upto $84/hr

mercor • Itália

Teletrabalho
EUR 178 223 000 - 215 096 000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Mercor • Lombardia

Presencial
EUR 90 000 - 130 000
Remote AI Safety Evaluator (Contract) — Expert Alignment & Risk
Remote AI Safety Evaluator (Contract) — Expert Alignment & Risk

mercor • Itália

Teletrabalho
EUR 74 000 - 86 000
AI Safety Evaluator & Model Alignment Specialist
AI Safety Evaluator & Model Alignment Specialist

mercor • Itália

Teletrabalho
EUR 74 000 - 86 000
Frontier AI Safety Evaluator
Frontier AI Safety Evaluator

Mercor • Roma

Teletrabalho
EUR 70 000 - 100 000
AI Safety Evaluator - Remote Contract
AI Safety Evaluator - Remote Contract

mercor • Itália

Teletrabalho
EUR 74 000 - 86 000
AI Safety Specialist - Fully Remote | Upto $70/hr
AI Safety Specialist - Fully Remote | Upto $70/hr

mercor • Itália

Teletrabalho
EUR 74 000 - 86 000