Get more replies from employers
Send a job-specific resume in minutes.
Mercor is seeking PhD and Master's scientists to author AI evaluation tasks for a new benchmark in scientific computing. You will create original, executable research problems and design grading criteria so frontier models struggle on these prompts.
The role involves sourcing data, writing prompts, and calibrating tasks against advanced models. 6 weeks, part-time (20+ hours/week), immediate start.
Mercor is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code)
Mercor is partnering with leading AI labs on a new benchmark for scientific computing. You will author original, executable research problems that today's frontier models cannot solve.