Founding AI Learning & Evaluation Scientist

Studyfetch

Beverly Hills (CA)

On-site

USD 150,000 - 210,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Medical, Dental, Vision (100% employer
75% dependent coverage
401(k) with employer matching
Daily team dinner in-office
Mission-driven small team

Job summary

Studyfetch Beverly Hills, CA, is seeking an AI Research Scientist focused on learning and evaluation. You will build the evaluation layer under model training and learning science across Learn Engine and the Honen platform, driving metrics that matter for student outcomes.

You’ll own multi-turn tutoring benchmarks, evaluate models in production, and align research with product decisions in a founding-team setting.

Qualifications

  • PhD in statistics, computer science, machine learning or quantitative field with 5+ years applying it to real products or research.
  • Experience evaluating AI products in production and a strong causal reasoning mindset.
  • Ability to design, run and defend experiments and benchmarks for AI models.
  • Strong communication skills and clarity in writing and speaking.

Responsibilities

  • Own the evaluation framework for the models we train and deployed.
  • Design and run evals against candidate models, including safety and reliability checks.
  • Build an internal benchmark for tutoring conversations and publish results when appropriate.
  • Bridge product data and model training to drive measurable improvements.

Skills

Evaluation for AI products
Experimental design
Causal inference
Bayesian methods
LLM evaluation

Education

PhD in statistics, CS, ML or related field

Tools

MongoDB
PostgreSQL
Vector databases
GCP

Job description

Studyfetch Beverly Hills, CA, is seeking an AI Research Scientist focused on learning and evaluation. You will build the evaluation layer under model training and learning science across Learn Engine and the Honen platform, driving metrics that matter for student outcomes.

You’ll own multi-turn tutoring benchmarks, evaluate models in production, and align research with product decisions in a founding-team setting.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Founding AI Learning & Evaluation Scientist — Equity
Founding AI Learning & Evaluation Scientist — Equity

Socket.dev • Beverly Hills (CA)

On-site
USD 180,000 - 280,000
Daily team dinner provided in-office
Staff Engineer, Applied AI for Education
Staff Engineer, Applied AI for Education

StudyFetch • Beverly Hills (CA)

On-site
USD 170,000 - 270,000
100% employer-paid Medical, Dental, and Vision
75% dependent coverage
401(k) with employer matching
+1
AI Research Scientist, Learning & Evaluation
AI Research Scientist, Learning & Evaluation

Studyfetch • Beverly Hills (CA)

On-site
USD 150,000 - 210,000
Medical, Dental, Vision (100% employer
75% dependent coverage
401(k) with employer matching
+2
AI Research Scientist, Learning & Evaluation
AI Research Scientist, Learning & Evaluation

Socket.dev • Beverly Hills (CA)

On-site
USD 180,000 - 280,000
Daily team dinner provided in-office
Chief AI Evaluation & Research
Chief AI Evaluation & Research

Vals AI, Inc. • San Francisco (CA)

On-site
USD 220,000 - 340,000
Relocation support
Health insurance
Meals provided (Lunch/Dinner)
+3
GenAI Evaluation Scientist - LLM Benchmarks & Failures
GenAI Evaluation Scientist - LLM Benchmarks & Failures

Scale • San Francisco (CA), Seattle (WA), New York (NY)

On-site
USD 166,000 - 207,000
Health insurance
Dental coverage
Vision coverage
+2
Head of AI Evaluation & Benchmarks
Head of AI Evaluation & Benchmarks

Vibehackers • San Francisco (CA), Northern (KY)

Hybrid
USD 225,000 - 275,000
Relocation and transportation support
Health and dental insurance
Lunch and dinner provided
+6
Senior Data Scientist, Education
Senior Data Scientist, Education

Learning Commons • Redwood City (CA)

On-site
USD 190,000 - 261,800
401(k) employer match
Paid volunteer time off
Relocation support
AI Tutor Research Scientist — Evaluation & Simulation
AI Tutor Research Scientist — Evaluation & Simulation

Heyaristotle • San Francisco (CA)

Hybrid
USD 85,000 - 120,000
Research, Post-Training Evals
Research, Post-Training Evals

Thinking Machines Lab Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 200,000