Get more replies from employers
Send a job-specific resume in minutes.
Mercor is seeking PhD/Master's scientists to author AI evaluation tasks and build benchmarks for scientific computing. You will craft original, executable research problems that current frontier models struggle with, sourcing material from papers, datasets, or open-source repositories.
Collaboration with AI labs is expected. Responsibilities include writing prompts, designing grading criteria, and calibrating models to ensure robust evaluation.
Mercor is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code)
Mercor is partnering with leading AI labs on a new benchmark for scientific computing. You will author original, executable research problems that today's frontier models cannot solve.