Frontier AI Safety & Alignment Evaluator

Obsidian

Roma

In loco

EUR 60.000 - 90.000

Tempo pieno

14 giorni+
Generatore di candidature

Una candidatura completa in un minuto — curriculum e lettera di presentazione personalizzati, pronti da inviare.

Supera i filtri ATS

Descrizione del lavoro

Obsidian is seeking experienced AI Safety Practitioners to evaluate safety, quality, and alignment of frontier AI models across sensitive, grey-area topics.

You will assess AI-generated responses, apply safety policies, and contribute to model improvement through structured evaluations and feedback.

Join a team collaborating with AI researchers to advance responsible AI safety benchmarking.

Competenze

  • Bachelor's degree or higher in relevant fields.
  • 5+ years of professional experience in AI Safety, Trust & Safety, journalism, public policy, or related fields.
  • Excellent written English and strong analytical reasoning.

Mansioni

  • Evaluate AI-generated responses for safety, factual accuracy, policy compliance, and overall quality.
  • Review content involving misinformation, political persuasion, self-harm, violence, cyber, biosecurity, and other sensitive domains.
  • Apply and refine evaluation rubrics for RLHF, SFT, and AI safety benchmarking.
  • Identify unsafe outputs, hallucinations, reasoning failures, and policy violations.
  • Provide structured feedback to improve model alignment and safety performance.
  • Collaborate with AI researchers and safety teams on ongoing evaluation initiatives.

Conoscenze

Excellent written English
Analytical reasoning
Critical thinking
Policy analysis

Formazione

Bachelor's degree or higher in journalism, communications, psychology, sociology, public policy, law, biology, chemistry, computer science, or related discipline

Descrizione del lavoro

Obsidian is seeking experienced AI Safety Practitioners to evaluate safety, quality, and alignment of frontier AI models across sensitive, grey-area topics.

You will assess AI-generated responses, apply safety policies, and contribute to model improvement through structured evaluations and feedback.

Join a team collaborating with AI researchers to advance responsible AI safety benchmarking.

Ottieni la revisione del curriculum gratis e riservata.
o trascina qui il file.
Similar jobs

Offerte di lavoro simili che vale la pena confrontare

Frontier AI Safety Evaluator & Benchmarking Specialist
Frontier AI Safety Evaluator & Benchmarking Specialist

Mercor • Milano

In loco
EUR 70.000 - 110.000
Frontier AI Safety Evaluator
Frontier AI Safety Evaluator

Mercor • Roma

Remoto
EUR 60.000 - 90.000
AI Safety Specialist - Evaluation Expert
AI Safety Specialist - Evaluation Expert

Obsidian • Roma

In loco
EUR 60.000 - 90.000
AI Safety Specialist - Evaluation Expert
AI Safety Specialist - Evaluation Expert

Mercor • Milano

In loco
EUR 70.000 - 110.000
Remote AI Safety Red Teamer - Adversarial Prompting
Remote AI Safety Red Teamer - Adversarial Prompting

Mercor • Roma

In loco
EUR 84.000 - 100.000
Remote AI Content Evaluator - Generalist Expert
Remote AI Content Evaluator - Generalist Expert

Synthires • Verona

Remoto
EUR 59.000 - 83.000
Fully remote
Remote AI Evaluation Specialist — UK/Europe
Remote AI Evaluation Specialist — UK/Europe

Synthires • Brescia

Remoto
EUR 59.000 - 83.000
Remote work opportunity
Flexible hours
Italian AI Safety Pro: Prompts & Moderation
Italian AI Safety Pro: Prompts & Moderation

Obsidian • Roma

In loco
EUR 28.000 - 38.000
AI Training Content Evaluator (UK/Europe)
AI Training Content Evaluator (UK/Europe)

Mercor • Roma

In loco
EUR 50.000 - 80.000
Remote AI Quality Evaluator (UK/Europe)
Remote AI Quality Evaluator (UK/Europe)

Synthires • Mantova

Remoto
EUR 59.000 - 83.000
Fully remote
Open to UK/Europe candidates