STEM PhD Expert for AI Reasoning & Evaluation

Braintrust

United States

Remote

AUD 173,000 - 288,000

Part time

10 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Braintrust is seeking PhD-level experts to help train and evaluate advanced AI models. This fully remote, flexible contract work shapes how models reason, answer, and improve.

You will review AI-generated content in your field, identify where the model is correct or incorrect, and help raise the quality bar on technical reasoning. The short-term engagement runs through the end of June, with potential to extend.

Qualifications

  • PhD completed or near completion in ML/AI, CS, engineering, statistics or related field.
  • Strong analytical and critical-thinking abilities to spot subtle errors in technical reasoning.
  • Fluent written English for clear technical communication.
  • Nice-to-have: research experience and data annotation or model evaluation.

Responsibilities

  • Assessing the factuality and relevance of domain-specific text produced by AI models.
  • Crafting and answering questions related to Machine Learning and AI.
  • Evaluating and ranking domain-specific responses generated by AI models.

Skills

Analytical thinking
Critical thinking
Fluent English writing

Education

PhD in ML/AI/CS/Engineering/Statistics or related

Job description

We're hiring PhD-level experts to help train and evaluate advanced AI models. This is remote, flexible contract work where your academic expertise directly shapes how cutting-edge models reason, answer, and improve. You'll review AI-generated content in your field, identify where the model gets things right or wrong, and help raise the quality bar on technical reasoning. This is a short-term engagement running through the end of June, with potential to extend depending on project needs.

Key Responsibilities
You May Contribute Your Expertise By
  • Assessing the factuality and relevance of domain-specific text produced by AI models
  • Crafting and answering questions related to Machine Learning and AI
  • Evaluating and ranking domain-specific responses generated by AI models
What We're Looking For
  • A PhD (completed or in final stages) in one of the following fields: Machine Learning / AI, Computer Science, Engineering, Statistics, or a closely related quantitative subdomain such as Mathematics or Physics
  • Strong analytical and critical-thinking skills, with the ability to spot subtle errors in technical reasoning
  • Fluent written English and the ability to communicate complex ideas clearly
Nice to Have
  • Research experience (academic or industry)
  • Prior experience with data annotation or AI model evaluation
  • Experience reviewing or publishing research papers
Compensation

Up to $150/hr, depending on your area of expertise, depth of experience, and assessment performance.

Location

This role is fully remote. We are currently accepting applicants based in:
United States, Canada, Puerto Rico, Mexico, United Kingdom, Australia, New Zealand, and Argentina.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

PhD-Level AI Model Evaluator & Domain Expert (Remote)
PhD-Level AI Model Evaluator & Domain Expert (Remote)

Braintrust • United States

Remote
AUD 173,000 - 288,000
AI Model Evaluation Scientist (PhD) — Remote Contract
AI Model Evaluation Scientist (PhD) — Remote Contract

Braintrust • Puerto Rico

On-site
USD 124,000 - 207,000
AI Research Expert (PhD) — Remote, Critical Analysis
AI Research Expert (PhD) — Remote, Critical Analysis

YO AI Labs • Austin (TX)

Remote
USD 90,000 - 130,000
AI Data Science Expert - Remote
AI Data Science Expert - Remote

YO AI Labs • Philadelphia

Remote
USD 40,000 - 75,000
Remote PhD Academic Expert for AI Research & Evaluation
Remote PhD Academic Expert for AI Research & Evaluation

YO AI Labs • Atlanta (GA)

Remote
USD 83,000 - 165,000
AI Data Scientist - Remote
AI Data Scientist - Remote

YO AI Labs • New York (NY)

Remote
USD 83,000 - 131,000
Remote PhD & Academic Expert for AI Research
Remote PhD & Academic Expert for AI Research

YO AI Labs • Phoenix (AZ)

Remote
USD 83,000 - 152,000
Remote PhD & Academic Expert — AI Research Evaluator
Remote PhD & Academic Expert — AI Research Evaluator

YO AI Labs • Washington

Remote
USD 83,000 - 124,000
AI Research Expert — Remote, PhD Level
AI Research Expert — Remote, PhD Level

YO AI Labs • Boston (MA)

Remote
USD 83,000 - 152,000
Remote PhD & Academic Expert for AI Research
Remote PhD & Academic Expert for AI Research

YO AI Labs • Dallas (TX)

Remote
USD 90,000 - 120,000