Remote Math AI Benchmark Designer (PhD)

turing

San Francisco (CA)

Hybrid

USD 138,000 - 207,000

Full time

47 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Fully remote
Flexible weekly hours

Job summary

Turing, based in San Francisco, California, seeks PhD candidates in Mathematics or Statistics for remote contract work to help fine-tune large language models. Earn $100+ per hour with fully remote, flexible weekly hours and no AI experience required.

You will design math problems, create benchmarks, and evaluate AI outputs, collaborating with researchers across advanced topics from undergraduate to PhD levels.

Qualifications

  • PhD (pursuing or completed) in Mathematics, Applied Math, Statistics, or related field.
  • Strong mathematical reasoning and problem-solving skills across advanced domains.
  • Ability to communicate complex ideas clearly in writing and provide structured feedback.
  • No AI experience required.

Responsibilities

  • Design advanced math problems to test AI performance (e.g., multi-step reasoning, abstraction, symbolic manipulation).
  • Develop clear, step-by-step solutions with rigorous logic.
  • Evaluate AI outputs for accuracy and quality of reasoning.
  • Collaborate with researchers to refine benchmarks across undergraduate to PhD-level math topics.

Skills

Mathematics
Analytical thinking
Academic writing
Benchmark design

Education

PhD in Mathematics
PhD in Statistics

Job description

Turing, based in San Francisco, California, seeks PhD candidates in Mathematics or Statistics for remote contract work to help fine-tune large language models. Earn $100+ per hour with fully remote, flexible weekly hours and no AI experience required.

You will design math problems, create benchmarks, and evaluate AI outputs, collaborating with researchers across advanced topics from undergraduate to PhD levels.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote Math Benchmark Designer for AI
Remote Math Benchmark Designer for AI

turing • San Francisco (CA)

Hybrid
USD 138,000 - 207,000
Fully remote
Leading LLM projects
Remote Math & AI Benchmark Designer (PhD)
Remote Math & AI Benchmark Designer (PhD)

Turing • San Francisco (CA)

Remote
USD 56,826 - 80,934
Remote Math & AI Benchmark Designer (PhD)
Remote Math & AI Benchmark Designer (PhD)

Turing • Chicago (IL)

Remote
USD 56,826 - 80,934
Remote Math & AI Benchmark Designer (PhD)
Remote Math & AI Benchmark Designer (PhD)

Turing • Toronto (OH)

Remote
USD 56,826 - 80,934
Remote Math Expert for AI Model Benchmarking
Remote Math Expert for AI Model Benchmarking

Parailabs • Northern (KY)

Hybrid
USD 90,000 - 130,000
Fully remote environment
Cutting-edge AI projects
Contractor engagement with potential续
Remote Mathematics Specialist
Remote Mathematics Specialist

turing • San Francisco (CA)

Hybrid
USD 138,000 - 207,000
Fully remote
Leading LLM projects
Remote Math Researcher - AI Benchmark & Problem Design
Remote Math Researcher - AI Benchmark & Problem Design

Turing • San Francisco (CA)

Hybrid
CAD 77,535 - 110,429
Remote Mathematics Specialist (PhD)
Remote Mathematics Specialist (PhD)

Turing • Chicago (IL)

On-site
USD 56,826 - 80,934
Remote Mathematics Specialist (PhD)
Remote Mathematics Specialist (PhD)

Turing • Toronto (OH)

On-site
USD 56,826 - 80,934
Remote Mathematics Researcher (PhD)
Remote Mathematics Researcher (PhD)

Turing • San Francisco (CA)

On-site
CAD 77,535 - 110,429