Mathematics PhD Coding Experts

Weekday 1

United States

On-site

USD 83,000 - 110,000

Part time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Weekday 1 is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code) for a new benchmark for scientific computing. You will author original, executable research problems that today's frontier models cannot solve.

You will source material, write scientific prompts, and build grading criteria. The role requires calibration against frontier models, with a 6-week, part-time commitment of 20+ hours per week and immediate start.

Qualifications

  • PhD in mathematics, applied mathematics, computational mathematics, or a closely related field.
  • Demonstrated depth in at least two of: numerical linear algebra, computational mechanics, computational finance.
  • Working proficiency in Python or R for scientific computing.
  • Comfortable with Git/GitHub and running code in Docker - authoring runs through a pull-request workflow with automated quality checks.

Responsibilities

  • Source your own material: a published paper, a Kaggle dataset, an open-source repository, or a scenario you design.
  • Write scientific prompts based on the input
  • Build the grading criteria that define a correct answer
  • Calibrate against frontier models - a task ships only when strong models fail it more often than they succeed

Skills

Python
R
Git/GitHub
Docker

Education

PhD in mathematics, applied mathematics, or computational mathematics

Tools

Docker
GitHub

Job description

Compensation: $70 per hour

We are hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code)

We are partnering with leading AI labs on a new benchmark for scientific computing. You will author original, executable research problems that today's frontier models cannot solve.

Domains - depth required in at least two subdomains (with a coding focus)
  • Mathematics - numerical linear algebra, computational mechanics, computational finance
Requirements
What you'll do
  • Source your own material: a published paper, a Kaggle dataset, an open-source repository, or a scenario you design
  • Write scientific prompts based on the input
  • Build the grading criteria that define a correct answer
  • Calibrate against frontier models - a task ships only when strong models fail it more often than they succeed
Required
  • PhD in mathematics, applied mathematics, computational mathematics, or a closely related field
  • Demonstrated depth in at least two of the following subdomains: numerical linear algebra, computational mechanics, computational finance
  • Working proficiency in Python or R for scientific computing
  • Comfortable with Git/GitHub and running code in Docker - authoring runs through a pull-request workflow with automated quality checks
Preferred
  • Publications in peer-reviewed journals
  • Prior scientific software or research engineering experience
Engagement
  • Duration: 6 weeks
  • Commitment: part-time, 20+ hours per week
  • Start date: immediate

We consider all qualified applicants without regard to legally protected characteristics and provide reasonable accommodations upon request.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Mathematics PhD Coding Experts
Mathematics PhD Coding Experts

RemoteLeads • Northern (KY)

Hybrid
USD 80,000 - 113,000
Six-week engagement
Accommodations available
Physics PhD Coding Experts
Physics PhD Coding Experts

Weekday 1 • United States

Remote
USD 83,000 - 110,000
Mathematics PhD - AI Evaluation Expert
Mathematics PhD - AI Evaluation Expert

Obsidian • San Francisco (CA)

On-site
USD 96,000 - 179,000
Biology PhD Coding Experts
Biology PhD Coding Experts

Weekday 1 • United States

Remote
USD 83,000 - 124,000
Material Science PhD Coding Experts
Material Science PhD Coding Experts

Weekday 1 • United States

Remote
USD 80,000 - 113,000
Mathematics PhD - AI Evaluation Expert - AI Trainer
Mathematics PhD - AI Evaluation Expert - AI Trainer

Obsidian • Miami (FL)

On-site
USD 83,000 - 124,000
Mathematics PhD - AI Evaluation Expert
Mathematics PhD - AI Evaluation Expert

Mercor • San Francisco (CA)

On-site
USD 11,021,000 - 13,776,000
Mathematics PhD - AI Evaluation Expert - AI Trainer
Mathematics PhD - AI Evaluation Expert - AI Trainer

Mercor • Los Angeles (CA)

On-site
USD 110,000 - 165,000
Mathematics PhD - AI Evaluation Expert - AI Trainer
Mathematics PhD - AI Evaluation Expert - AI Trainer

Obsidian • Los Angeles (CA)

On-site
USD 83,000 - 124,000
Mathematics PhD - AI Evaluation Expert - AI Trainer
Mathematics PhD - AI Evaluation Expert - AI Trainer

Mercor • Miami (FL)

On-site
USD 83,000 - 138,000