An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Mercor is seeking PhD/Master's scientists to author AI evaluation tasks and build benchmarks for scientific computing. You will craft original, executable research problems that current frontier models struggle with, sourcing material from papers, datasets, or open-source repositories.
Collaboration with AI labs is expected. Responsibilities include writing prompts, designing grading criteria, and calibrating models to ensure robust evaluation.
Mercor is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code)
Mercor is partnering with leading AI labs on a new benchmark for scientific computing. You will author original, executable research problems that today's frontier models cannot solve.