AI Model Evaluation Scientist (PhD) — Remote Contract

Braintrust

Puerto Rico

On-site

USD 124,000 - 207,000

Part time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Braintrust is seeking PhD-level experts to help train and evaluate advanced AI models. This fully remote, flexible contract role lets you apply domain expertise to shape how models reason, answer, and improve.

You will review AI-generated content in your field, identify where the model is right or wrong, and raise the quality bar on technical reasoning. The engagement runs through the end of June, with potential to extend depending on project needs.

Qualifications

  • PhD in ML/AI/CS/Engineering/Statistics or related field.
  • Strong analytical and critical-thinking abilities.
  • Fluent written English with clear communication.

Responsibilities

  • Assessing the factuality and relevance of domain-specific text produced by AI models
  • Crafting and answering questions related to Machine Learning and AI
  • Evaluating and ranking domain-specific responses generated by AI models

Skills

Analytical thinking
Critical thinking
Fluent English

Education

PhD in ML/AI/CS/Engineering/Statistics or related field

Job description

Braintrust is seeking PhD-level experts to help train and evaluate advanced AI models. This fully remote, flexible contract role lets you apply domain expertise to shape how models reason, answer, and improve.

You will review AI-generated content in your field, identify where the model is right or wrong, and raise the quality bar on technical reasoning. The engagement runs through the end of June, with potential to extend depending on project needs.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

PhD-Level AI Model Evaluator & Domain Expert (Remote)
PhD-Level AI Model Evaluator & Domain Expert (Remote)

Braintrust • United States

Remote
AUD 173,000 - 288,000
STEM PhD Expert for AI Reasoning & Evaluation
STEM PhD Expert for AI Reasoning & Evaluation

Braintrust • United States

Remote
AUD 173,000 - 288,000
Remote AI Finance Research Evaluator (Contract)
Remote AI Finance Research Evaluator (Contract)

Braintrust • Puerto Rico

On-site
USD 90,000 - 130,000
Remote work
Short-term engagement through June
AI Research Expert (PhD) — Remote, Critical Analysis
AI Research Expert (PhD) — Remote, Critical Analysis

YO AI Labs • Austin (TX)

Remote
USD 90,000 - 130,000
AI Research Quality Evaluator — Remote (PhD Expert)
AI Research Quality Evaluator — Remote (PhD Expert)

YO AI Labs • California City (CA)

Remote
USD 96,000 - 165,000
AI Research Quality Analyst (Remote, PhD)
AI Research Quality Analyst (Remote, PhD)

YO AI Labs • Washington

Remote
USD 83,000 - 165,000
AI Research Evaluator & Subject-Matter Expert Remote
AI Research Evaluator & Subject-Matter Expert Remote

YO AI Labs • Los Angeles (CA)

Remote
USD 83,000 - 165,000
AI Researcher & Academic Expert (PhD) — Remote (Contract)
AI Researcher & Academic Expert (PhD) — Remote (Contract)

YO AI Labs • Houston (TX)

Remote
USD 90,000 - 130,000
AI Research Expert — Remote, PhD Level
AI Research Expert — Remote, PhD Level

YO AI Labs • Boston (MA)

Remote
USD 83,000 - 152,000
PhD AI Research Reviewer (Remote Contractor)
PhD AI Research Reviewer (Remote Contractor)

YO AI Labs • Maryland City (MD)

Remote
USD 83,000 - 124,000