Remote AI Safety Red Team Specialist

Mercor

Lisboa

Presencial

EUR 40 000 - 65 000

Tempo integral

14 dias+
Gerador de candidaturas

Destaca-te para esta função — gera um currículo e uma carta de apresentação personalizados em cerca de um minuto.

Ultrapassa os filtros ATS

Resumo da oferta

Mercor is assembling a red team to test AI models with adversarial inputs, surface vulnerabilities, and generate data to improve safety for customers. The role involves reviewing outputs on sensitive topics and collaborating with stakeholders across projects.

Participation in high-sensitivity tasks is supported with guidelines and wellness resources. You will contribute by red-teaming models, generating high-quality data, following taxonomies, and documenting reproducible reports and attack

Qualificações

  • Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing.
  • Must be curious and adversarial, able to push systems to breaking points.
  • Structured approach using frameworks or benchmarks, not ad-hoc hacks.
  • Clear communication of risks to technical and non-technical stakeholders.
  • Able to adapt across projects and customers.

Responsabilidades

  • Red team conversational AI models and agents: jailbreaks, prompt injections, misuse cases, bias exploitation, multi-turn manipulation.
  • Generate high-quality human data: annotate failures, classify vulnerabilities, and flag systemic risks.
  • Apply structure: follow taxonomies, benchmarks, and playbooks to keep testing consistent.
  • Document reproducibly: produce reports, datasets, and attack cases customers can act on.

Conhecimentos

Red teaming experience
Adversarial mindset
Structured thinking
Clear communication
Adaptability

Descrição da oferta de emprego

Mercor is assembling a red team to test AI models with adversarial inputs, surface vulnerabilities, and generate data to improve safety for customers. The role involves reviewing outputs on sensitive topics and collaborating with stakeholders across projects.

Participation in high-sensitivity tasks is supported with guidelines and wellness resources. You will contribute by red-teaming models, generating high-quality data, following taxonomies, and documenting reproducible reports and attack

Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Remote AI Safety Red Teamer
Remote AI Safety Red Teamer

Mercor • Lisboa

Teletrabalho
EUR 90 000 - 130 000
Remote AI Safety Red Team Specialist (English/Portuguese)
Remote AI Safety Red Team Specialist (English/Portuguese)

Mercor • Lisboa

Presencial
EUR 70 000 - 110 000
Remote AI Safety Red Teamer - Stress-Test Frontier Models
Remote AI Safety Red Teamer - Stress-Test Frontier Models

Mercor • Lisboa

Presencial
EUR 83 000 - 100 000
Elite Frontier AI Safety Red Teamer
Elite Frontier AI Safety Red Teamer

Mercor • Lisboa

Presencial
EUR 60 000 - 90 000
AI Safety Red Teamer — Remote
AI Safety Red Teamer — Remote

Mercor • Lisboa

Presencial
EUR 69 000 - 104 000
Bilingual AI Safety Red Teamer (Remote)
Bilingual AI Safety Red Teamer (Remote)

Mercor • Lisboa

Presencial
EUR 45 000 - 65 000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Mercor • Lisboa

Presencial
EUR 60 000 - 90 000
AI Safety Red Team Specialist — Remote (EN/PT)
AI Safety Red Team Specialist — Remote (EN/PT)

Crossing Hurdles • Portugal

Presencial
EUR 35 000 - 54 000
Frontier AI Safety Evaluator & Policy Alignment Expert
Frontier AI Safety Evaluator & Policy Alignment Expert

Mercor • Lisboa

Presencial
EUR 65 000 - 100 000
AI Safety Evaluator & Model Alignment Specialist
AI Safety Evaluator & Model Alignment Specialist

Mercor • Lisboa

Teletrabalho
EUR 60 000 - 90 000