Math PhD Computational Scientist – Part-Time AI Benchmarking

RemoteLeads

Northern (KY)

Hybrid

USD 80,000 - 113,000

Part time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Six-week engagement
Accommodations available

Job summary

RemoteLeads seeks a highly qualified researcher to work part-time for six weeks at 20+ hours per week, starting immediately. The role focuses on creating scientific computing prompts, grading criteria, and benchmarking AI models against frontier challenges.

Proficiency in Python or R is required, with a PhD in a relevant field and strong computational background. Candidates should be comfortable with Git and Docker, and prior research engineering or peer-reviewed publications are a plus.

Qualifications

  • PhD in mathematics, applied mathematics, computational mathematics, or closely related field.
  • Depth in at least two of numerical linear algebra, computational mechanics, and computational finance.
  • Working proficiency in Python or R for scientific computing.
  • Comfort with Git/GitHub and running code in Docker.
  • Peer‑reviewed journal publications preferred.
  • Prior scientific software or research engineering experience preferred.
  • Availability for part-time work of 20+ hours per week over six weeks.
  • Ability to start immediately.

Responsibilities

  • Source material from research papers, Kaggle datasets, open-source repositories, or original scenarios.
  • Write scientific computing prompts based on the selected material.
  • Develop grading criteria that define correct solutions.
  • Calibrate tasks against frontier AI models and release them only when models frequently fail.
  • Submit authored work through a GitHub pull-request workflow with automated quality checks.

Skills

Python
R
GitHub
Docker

Education

PhD in mathematics / applied mathematics / computational mathematics

Tools

Git
Docker

Job description

RemoteLeads seeks a highly qualified researcher to work part-time for six weeks at 20+ hours per week, starting immediately. The role focuses on creating scientific computing prompts, grading criteria, and benchmarking AI models against frontier challenges.

Proficiency in Python or R is required, with a PhD in a relevant field and strong computational background. Candidates should be comfortable with Git and Docker, and prior research engineering or peer-reviewed publications are a plus.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Mathematics PhD: Scientific AI Prompt Engineer (Part-Time)
Mathematics PhD: Scientific AI Prompt Engineer (Part-Time)

RemoteLeads • United States

Remote
USD 83,000 - 117,000
Competitive pay at $70 per hour
Six-week part-time engagement
Opportunity to contribute to AIbench
+1
Bio PhD AI Benchmark Architect (Part-Time, 6-Week Project)
Bio PhD AI Benchmark Architect (Part-Time, 6-Week Project)

Mercor • San Diego (CA)

On-site
USD 69,000 - 103,000
Remote AI Benchmark Scientist (Physics PhD)
Remote AI Benchmark Scientist (Physics PhD)

Weekday 1 • United States

Remote
USD 83,000 - 110,000
AI Benchmark Scientist - Mathematics PhD (6-Week Project)
AI Benchmark Scientist - Mathematics PhD (6-Week Project)

Weekday 1 • United States

On-site
USD 83,000 - 110,000
Chemistry PhD AI Benchmark Architect (Remote)
Chemistry PhD AI Benchmark Architect (Remote)

Weekday 1 • United States

Remote
USD 83,000 - 110,000
Remote Data Science & AI Benchmark Analyst
Remote Data Science & AI Benchmark Analyst

Weekday 1 • United States

Remote
USD 109,000 - 164,000
Fully remote
Weekly payments
Remote Bio AI Evaluation Scientist (PhD)
Remote Bio AI Evaluation Scientist (PhD)

Weekday 1 • United States

Remote
USD 83,000 - 124,000
AI Evaluation Scientist - Math PhD (6-Week, Part-Time)
AI Evaluation Scientist - Math PhD (6-Week, Part-Time)

Mercor • Miami (FL)

On-site
USD 83,000 - 138,000
6-Week Part-Time Materials Scientist for AI Benchmarks
6-Week Part-Time Materials Scientist for AI Benchmarks

Mercor • San Francisco (CA)

Remote
USD 150,000 - 190,000
Quantum Computing AI Benchmark Scientist (PhD) — Part-Time
Quantum Computing AI Benchmark Scientist (PhD) — Part-Time

Mercor • New York (NY)

On-site
USD 83,000 - 124,000