Lead, AI Safety Benchmarks & Evaluations

Alice (Formerly ActiveFence)

United States

On-site

USD 180,000 - 280,000

Full time

11 days ago
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Alice (Formerly ActiveFence) is seeking a senior leader to drive safety and evaluation research for frontier AI models. You will own the taxonomy, oversee the benchmark suite, and coordinate a small pool of freelancers while aligning with the CTO's public research agenda.

You will work with 150 researchers, travel to conferences, and shape quarterly release plans based on client asks and emerging events in AI safety and security.

Qualifications

  • PhD or Masters in computer science, machine learning or a related field, or equivalent depth from industry research.
  • 3+ years building and running safety or security evaluations for language models in production or research settings.
  • 5+ relevant research publications in AI safety and security including lead author on at least 2 of them.
  • Strong engineering background: evaluation harnesses, distributed inference, vLLM, and reading/fixing codebases.
  • You can build a taxonomy, not only score against one.

Responsibilities

  • Ship the benchmark cadence and manage the schedule for frontiers evaluation cycles.
  • Own the quality bar: ensure clear rubrics, sane distributions, and taxonomy alignment.
  • Run the process: hold timelines, coordinate researchers, and direct freelancers when needed.
  • Set the roadmap with the forum: collaborate with CTO, pod, and leads to refresh quarterly plans.
  • Stay ahead of the curve: engage with labs, researchers, and conferences; weekly contact.

Job description

Alice (Formerly ActiveFence) is seeking a senior leader to drive safety and evaluation research for frontier AI models. You will own the taxonomy, oversee the benchmark suite, and coordinate a small pool of freelancers while aligning with the CTO's public research agenda.

You will work with 150 researchers, travel to conferences, and shape quarterly release plans based on client asks and emerging events in AI safety and security.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Safety Benchmark Lead & Evaluation Architect
AI Safety Benchmark Lead & Evaluation Architect

Alice • New York (NY)

On-site
USD 180,000 - 250,000
AI Safety Benchmarks Lead — Ship Frontiers & Roadmaps
AI Safety Benchmarks Lead — Ship Frontiers & Roadmaps

Alice (Formerly ActiveFence) • New York (NY)

On-site
USD 190,000 - 240,000
Frontier AI Safety Benchmark Lead
Frontier AI Safety Benchmark Lead

Alice • San Francisco (CA)

Hybrid
USD 180,000 - 230,000
Conference travel
Research Lead, Evaluations and Benchmarks
Research Lead, Evaluations and Benchmarks

Alice (Formerly ActiveFence) • New York (NY)

On-site
USD 190,000 - 240,000
Research Lead, Evaluations and Benchmarks
Research Lead, Evaluations and Benchmarks

Alice (Formerly ActiveFence) • United States

On-site
USD 180,000 - 280,000
Research Lead: AI Safety & Pre-Training at Scale
Research Lead: AI Safety & Pre-Training at Scale

FAR.AI • Berkeley (CA)

On-site
USD 180,000 - 240,000
Research Lead, AI Safety & Impact
Research Lead, AI Safety & Impact

FAR.AI • United States

Hybrid
USD 170,000 - 270,000
Catered lunch and dinner at Berkeley
Lead Frontier AI Safety Red Teamer
Lead Frontier AI Safety Red Teamer

Obsidian • San Francisco (CA)

On-site
USD 150,000 - 230,000
Frontier AI Safety Evaluator & Alignment Expert
Frontier AI Safety Evaluator & Alignment Expert

Obsidian • New York (NY)

On-site
USD 120,000 - 180,000
Member of the Technical Staff
Member of the Technical Staff

Alice (Formerly ActiveFence) • San Francisco (CA)

On-site
USD 170,000 - 260,000
Bay Area base
Public speaking opportunities