Remote Math Expert for AI Benchmarking & Evaluation

Turing

Lineville (IA)

Remote

USD 114,000 - 162,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Fully remote
Flexible weekly hours
Competitive pay at $100/hour
Contract with defined start/end dates
Up to 40 hrs/week

Job summary

Turing is seeking PhD-level mathematicians to design challenging AI evaluation problems and craft rigorous, step-by-step solutions. You will assess AI reasoning, ensure clarity of feedback, and collaborate with researchers to build robust benchmarks across math topics from undergraduate to PhD level.

This fully remote contract role offers flexible hours and a $100/hour pay rate. Candidates should have a PhD in mathematics or related fields, strong reasoning, and communication skills.

Qualifications

  • PhD (pursuing or completed) in Mathematics, Applied Math, Statistics, or related field.
  • Strong mathematical reasoning and problem solving across advanced domains.
  • Ability to communicate complex ideas clearly in writing and provide feedback.
  • No AI experience required.

Responsibilities

  • Design advanced math problems to test AI performance (e.g., multi-step reasoning, abstraction, symbolic manipulation).
  • Develop clear, step-by-step solutions with rigorous logic.
  • Evaluate AI outputs for accuracy and quality of reasoning.
  • Collaborate with researchers to refine benchmarks across undergraduate to PhD-level math topics.

Skills

Mathematical reasoning
Problem solving
Written communication
Attention to detail

Education

PhD in Mathematics/Applied Math/Statistics

Job description

Turing is seeking PhD-level mathematicians to design challenging AI evaluation problems and craft rigorous, step-by-step solutions. You will assess AI reasoning, ensure clarity of feedback, and collaborate with researchers to build robust benchmarks across math topics from undergraduate to PhD level.

This fully remote contract role offers flexible hours and a $100/hour pay rate. Candidates should have a PhD in mathematics or related fields, strong reasoning, and communication skills.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote Math & AI Benchmark Designer (PhD)
Remote Math & AI Benchmark Designer (PhD)

Turing • San Francisco (CA)

Remote
USD 80,000 - 100,000
Remote Math & AI Benchmark Designer (PhD)
Remote Math & AI Benchmark Designer (PhD)

Turing • Chicago (IL)

Remote
Remote Math Researcher - AI Benchmark & Problem Design
Remote Math Researcher - AI Benchmark & Problem Design

Turing • San Francisco (CA)

Hybrid
Remote Mathematics Specialist (PhD)
Remote Mathematics Specialist (PhD)

Turing • San Francisco (CA)

Hybrid
Remote Math & Stats PhD: AI Benchmark Designer
Remote Math & Stats PhD: AI Benchmark Designer

turing • San Francisco (CA)

Hybrid
USD 138,000 - 207,000
Fully remote
Flexible weekly hours
Remote Math & AI Benchmark Designer (PhD)
Remote Math & AI Benchmark Designer (PhD)

Turing • Toronto (OH)

Remote
USD 80,000 - 100,000
Remote Mathematics Specialist
Remote Mathematics Specialist

turing • San Francisco (CA)

Hybrid
USD 138,000 - 207,000
Fully remote
Flexible weekly hours
Remote Applied Mathematician for AI Benchmarking
Remote Applied Mathematician for AI Benchmarking

24-Mag Llc • New York (NY)

Remote
USD 69,000 - 96,000
Remote work
Part-time contract
Flexible scheduling
Remote Mathematics Specialist (PhD)
Remote Mathematics Specialist (PhD)

Turing • Toronto (OH)

Remote
Remote Mathematics Specialist (PhD)
Remote Mathematics Specialist (PhD)

Turing • Chicago (IL)

Remote