A complete application in a minute — tailored resume and cover letter, ready to send.
Weekday 1 is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code) for a new benchmark for scientific computing. You will author original, executable research problems that today's frontier models cannot solve.
You will source material, write scientific prompts, and build grading criteria. The role requires calibration against frontier models, with a 6-week, part-time commitment of 20+ hours per week and immediate start.
We are hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code)
We are partnering with leading AI labs on a new benchmark for scientific computing. You will author original, executable research problems that today's frontier models cannot solve.
We consider all qualified applicants without regard to legally protected characteristics and provide reasonable accommodations upon request.