AI Evaluation Benchmark Fellow

Snorkel AI

New York, San Francisco (NY, CA)

Hybrid

USD 117,000 - 227,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Research advisor
Compute resources
Open-source artifact
Community events

Job summary

Snorkel AI is seeking a Research Fellow to advance open benchmarks and evaluation environments that shape how AI progress is measured. You will own the project from research question to public release, with a Snorkel co-advisor and access to compute and domain experts.

You’ll spend time between your academic affiliation and Snorkel’s research team, with opportunities to work in our San Francisco or New York offices where applicable.

Qualifications

  • PhD candidates or postdocs with a strong research background in AI, evaluation, or related fields
  • Ability to turn research ideas into working code, datasets, or tools
  • Experience collaborating with domain experts to turn judgments into specifications
  • Capacity to scope a project to a public release within the fellowship term

Responsibilities

  • Scope and lead a new benchmark, dataset, or environment in a domain where current evaluations fall short
  • Define the task specification, rubrics, and reference answers with domain experts from Snorkel's network, and run the expert review process
  • Release the research artifact publicly, with you as lead author
  • Join weekly research meetings to share progress, receive feedback, and contribute to other researchers’ projects

Job description

Snorkel AI is seeking a Research Fellow to advance open benchmarks and evaluation environments that shape how AI progress is measured. You will own the project from research question to public release, with a Snorkel co-advisor and access to compute and domain experts.

You’ll spend time between your academic affiliation and Snorkel’s research team, with opportunities to work in our San Francisco or New York offices where applicable.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Fellow: Build Open AI Benchmarks & Evaluations
Research Fellow: Build Open AI Benchmarks & Evaluations

Snorkel AI • United States

Hybrid
USD 117,000 - 227,000
Co-advisor matched to project
Compute resources for your project
Open-source research artifact release
+1
Research Fellowship
Research Fellowship

Snorkel AI • United States

Hybrid
USD 117,000 - 227,000
Co-advisor matched to project
Compute resources for your project
Open-source research artifact release
+1
Research Fellowship
Research Fellowship

Snorkel AI • New York (NY), San Francisco (CA)

Hybrid
USD 117,000 - 227,000
Research advisor
Compute resources
Open-source artifact
+1
Frontier AI Benchmarks Scientist (Data-Centric)
Frontier AI Benchmarks Scientist (Data-Centric)

Snorkel AI • San Francisco (CA)

Hybrid
USD 200,000 - 325,000
Head of Research: Data Evaluation & Training
Head of Research: Data Evaluation & Training

Snorkel AI • New York (NY)

On-site
USD 275,000 - 425,000
Research Scientist — Frontier Data Benchmarks & AI Evaluation
Research Scientist — Frontier Data Benchmarks & AI Evaluation

Snorkel AI • United States

On-site
USD 200,000 - 350,000
Head of AI Data Evaluation & Research
Head of AI Data Evaluation & Research

Snorkel AI • United States

On-site
USD 275,000 - 425,000
Frontier AI Benchmark Scientist
Frontier AI Benchmark Scientist

Snorkel AI • New York (NY)

On-site
USD 200,000 - 350,000
Director, Research - Evaluation & Training
Director, Research - Evaluation & Training

Snorkel AI • New York (NY)

On-site
USD 275,000 - 425,000
Research Scientist - Frontier Benchmarks
Research Scientist - Frontier Benchmarks

Snorkel AI • New York (NY)

On-site
USD 200,000 - 350,000