Frontier AI Safety Evaluator & Policy Alignment Expert

Mercor

Lisboa

Presencial

EUR 65 000 - 100 000

Tempo integral

14 dias+

Recebe mais respostas dos empregadores

Envia um currículo específico para a oferta em poucos minutos.

Resumo da oferta

Mercor is seeking experienced AI Safety Practitioners to evaluate frontier AI models across safety, alignment, and policy-sensitive topics. You will assess AI-generated responses, apply safety policies, and provide structured feedback to improve model behavior.

Collaborate with researchers and safety teams on RLHF/SFT evaluations, review misinformation, self-harm, violence, and other sensitive domains, and help shape safer AI used by millions.

Qualificações

  • Bachelor's degree or higher in a related discipline.
  • 5+ years of professional experience in AI Safety, Trust & Safety, journalism, public policy, scientific research, security, or a related field.
  • Excellent written English, critical thinking, and analytical reasoning skills.
  • Ability to consistently evaluate nuanced and policy-sensitive scenarios.

Responsabilidades

  • Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality.
  • Review content involving misinformation, political persuasion, self-harm, violence, cyber, biosecurity, and other sensitive domains.
  • Apply and refine evaluation rubrics for RLHF, SFT, and AI safety benchmarking.
  • Identify unsafe outputs, hallucinations, reasoning failures, and policy violations.
  • Provide structured feedback to improve model alignment and safety performance.
  • Collaborate with AI researchers and safety teams on ongoing evaluation initiatives.

Conhecimentos

Excellent written English
Critical thinking
Analytical reasoning
Policy evaluation

Formação académica

Bachelor's degree or higher in a related discipline

Descrição da oferta de emprego

Mercor is seeking experienced AI Safety Practitioners to evaluate frontier AI models across safety, alignment, and policy-sensitive topics. You will assess AI-generated responses, apply safety policies, and provide structured feedback to improve model behavior.

Collaborate with researchers and safety teams on RLHF/SFT evaluations, review misinformation, self-harm, violence, and other sensitive domains, and help shape safer AI used by millions.

Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Frontier AI Safety Evaluator
Frontier AI Safety Evaluator

Mercor • Lisboa

Teletrabalho
EUR 45 000 - 65 000
Frontier AI Safety Evaluator & Policy Auditor
Frontier AI Safety Evaluator & Policy Auditor

Obsidian • Lisboa

Presencial
EUR 55 000 - 90 000
AI Safety Practitioner - Expert Evaluator
AI Safety Practitioner - Expert Evaluator

Mercor • Lisboa

Presencial
EUR 65 000 - 100 000
AI Safety Specialist - Evaluation Expert
AI Safety Specialist - Evaluation Expert

Mercor • Lisboa

Presencial
EUR 60 000 - 90 000
AI Safety Specialist - Evaluation Expert
AI Safety Specialist - Evaluation Expert

Obsidian • Lisboa

Presencial
EUR 55 000 - 90 000
Elite Frontier AI Safety Red Teamer
Elite Frontier AI Safety Red Teamer

Mercor • Lisboa

Presencial
EUR 60 000 - 90 000
Remote AI Safety Red Teamer - Stress-Test Frontier Models
Remote AI Safety Red Teamer - Stress-Test Frontier Models

Mercor • Lisboa

Presencial
EUR 83 000 - 100 000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • Lisboa

Presencial
EUR 40 000 - 65 000
AI Safety Red Teamer Expert
AI Safety Red Teamer Expert

Mercor • Lisboa

Presencial
EUR 60 000 - 90 000
Senior Regulatory Counsel - AI Compliance & Global Policy
Senior Regulatory Counsel - AI Compliance & Global Policy

Obsidian • Lisboa

Teletrabalho
EUR 70 000 - 120 000