An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Mercor is seeking PhD and Master’s scientists to author AI evaluation tasks for a new benchmark in scientific computing. You will source material, craft executable problems, and build the grading criteria to measure model performance on challenging scenarios.
You will calibrate prompts against frontier models, ensuring tasks remain difficult where state-of-the-art systems struggle, across two subdomains with a coding focus.
Mercor is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code)
Mercor is partnering with leading AI labs on a new benchmark for scientific computing. You will author original, executable research problems that today's frontier models cannot solve.