Remote Math & Stats PhD: AI Benchmark Designer

turing

San Francisco (CA)

Hybrid

USD 138,000 - 207,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Fully remote
Flexible weekly hours

Job summary

Turing is seeking PhDs in Mathematics or Statistics for a fully remote contract to help fine-tune large language models. You will design problems, test AI solutions, and collaborate on benchmarks with leading AI labs.

Requirements include a PhD (pursuing or completed) in mathematics or related fields, strong mathematical reasoning, and clear writing. No AI experience required; up to 40 hours per week with flexible scheduling.

Qualifications

  • PhD in Mathematics or related field required or in progress.
  • Strong mathematical reasoning and problem-solving skills across advanced domains.
  • Ability to communicate complex ideas clearly in writing and provide structured feedback.
  • No AI experience required.

Responsibilities

  • Design advanced math problems to test AI performance (multi-step reasoning, abstraction, symbolic manipulation).
  • Develop clear, step-by-step solutions with rigorous logic.
  • Evaluate AI outputs for accuracy and quality of reasoning.
  • Collaborate with researchers to refine benchmarks across math topics from undergraduate to PhD level.

Skills

PhD in Mathematics
Strong mathematical reasoning
Analytical skills
Clear written communication

Education

PhD (pursuing or completed) in Mathematics, Applied Math, Statistics, or related field

Job description

Turing is seeking PhDs in Mathematics or Statistics for a fully remote contract to help fine-tune large language models. You will design problems, test AI solutions, and collaborate on benchmarks with leading AI labs.

Requirements include a PhD (pursuing or completed) in mathematics or related fields, strong mathematical reasoning, and clear writing. No AI experience required; up to 40 hours per week with flexible scheduling.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote Math PhD — AI Benchmark Design & Evaluation
Remote Math PhD — AI Benchmark Design & Evaluation

United States Digital Space LLC • United States

Remote
Fully remote work
Flexible hours
Cutting-edge AI projects
Remote Math PhD — Design AI Benchmark Problems
Remote Math PhD — Design AI Benchmark Problems

United States Digital Space LLC • United States

Remote
Fully remote
Flexible weekly hours
Work with leading AI labs
Remote Math Expert for AI Benchmarking & Evaluation
Remote Math Expert for AI Benchmarking & Evaluation

Turing • Lineville (IA)

Remote
USD 114,000 - 162,000
Fully remote
Flexible weekly hours
Competitive pay at $100/hour
+2
Remote Mathematics Benchmark Architect for AI
Remote Mathematics Benchmark Architect for AI

Turing • San Francisco (CA)

Remote
USD 138,000 - 207,000
Fully remote
Cutting-edge AI projects
Remote Math Researcher for AI Benchmarks
Remote Math Researcher for AI Benchmarks

United States Digital Space LLC • United States

Remote
Fully remote, flexible work
Work on cutting-edge AI projects
Remote Math & AI Benchmark Designer (PhD)
Remote Math & AI Benchmark Designer (PhD)

Turing • Chicago (IL)

Remote
Remote Math & AI Benchmark Designer (PhD)
Remote Math & AI Benchmark Designer (PhD)

Turing • San Francisco (CA)

Remote
USD 80,000 - 100,000
Remote Mathematics Specialist
Remote Mathematics Specialist

turing • San Francisco (CA)

Hybrid
USD 138,000 - 207,000
Fully remote
Flexible weekly hours
Remote Math & AI Benchmark Designer (PhD)
Remote Math & AI Benchmark Designer (PhD)

Turing • Toronto (OH)

Remote
USD 80,000 - 100,000
Remote Mathematics Specialist (PhD)
Remote Mathematics Specialist (PhD)

Turing • Chicago (IL)

Remote