Frontier AI Safety Evaluator

Mercor

Roma

Remoto

EUR 70.000 - 100.000

Tempo pieno

14 giorni+
Generatore di candidature

Ottieni una risposta da questo datore di lavoro — un curriculum e una lettera di presentazione personalizzati, che corrispondono esattamente a ciò che sta cercando.

Supera i filtri ATS

Descrizione del lavoro

Mercor is seeking experienced AI Safety Practitioners to evaluate safety, quality, and alignment of frontier AI models across complex, policy-sensitive topics. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations and feedback.

The role involves reviewing content on misinformation, political persuasion, self-harm, violence, biosecurity, and other sensitive domains; applying and refining evaluation rubrics for RLHF, SFT, and

Competenze

  • Bachelor's degree or higher in a related field.
  • 5+ years of professional experience in AI Safety, Trust & Safety, journalism, public policy, or related field.
  • Excellent written English and analytical reasoning.
  • Ability to evaluate nuanced and policy-sensitive scenarios.

Mansioni

  • Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and quality.
  • Review content involving misinformation, political persuasion, self-harm, violence, and other sensitive domains.
  • Apply and refine evaluation rubrics for RLHF and SFT.
  • Identify unsafe outputs, hallucinations, and policy violations.
  • Provide structured feedback to improve model alignment and safety performance.
  • Collaborate with AI researchers and safety teams on evaluation initiatives.

Conoscenze

Excellent written English
Critical thinking
Analytical reasoning
Nuanced content evaluation

Formazione

Bachelor's degree or higher in a related field

Descrizione del lavoro

Mercor is seeking experienced AI Safety Practitioners to evaluate safety, quality, and alignment of frontier AI models across complex, policy-sensitive topics. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations and feedback.

The role involves reviewing content on misinformation, political persuasion, self-harm, violence, biosecurity, and other sensitive domains; applying and refining evaluation rubrics for RLHF, SFT, and

Ottieni la revisione del curriculum gratis e riservata.
o trascina qui il file.
Similar jobs

Offerte di lavoro simili che vale la pena confrontare

Frontier AI Safety Evaluator: Expert Feedback
Frontier AI Safety Evaluator: Expert Feedback

Mercor • Roma

In loco
EUR 65.000 - 95.000
Frontier AI Content Quality Assessor
Frontier AI Content Quality Assessor

Mercor • Roma

In loco
EUR 45.000 - 75.000
Frontier AI Safety Red Teamer - Adversarial Testing Lead
Frontier AI Safety Red Teamer - Adversarial Testing Lead

Mercor • Milano

In loco
EUR 90.000 - 130.000
Remote AI Safety Red Teamer - Adversarial Prompting
Remote AI Safety Red Teamer - Adversarial Prompting

Mercor • Roma

In loco
EUR 84.000 - 100.000
Remote AI Ethics & Safety Analyst
Remote AI Ethics & Safety Analyst

Alignerr • Italia

In loco
EUR 24.000 - 72.000
Fully remote
Flexible schedule
Freelance autonomy
+2
AI Training Content Evaluator (UK/Europe)
AI Training Content Evaluator (UK/Europe)

Mercor • Roma

In loco
EUR 50.000 - 80.000
Explosives & Energetic Materials Expert for AI Safety Audit
Explosives & Energetic Materials Expert for AI Safety Audit

SME Careers • Italia

In loco
EUR 70.000 - 120.000
Remote AI Safety Tester
Remote AI Safety Tester

TSMG Holding • Roma

Remoto
EUR 26.000 - 39.000
Remote Medical AI Quality & Safety Reviewer
Remote Medical AI Quality & Safety Reviewer

Turing • Italia

Remoto
EUR 40.000 - 60.000
Generalist Expert - Content Evaluator
Generalist Expert - Content Evaluator

Mercor • Roma

In loco
EUR 45.000 - 75.000