A complete application in a minute — tailored resume and cover letter, ready to send.
Mercor is seeking PhD- and Master’s-level scientists to author AI evaluation tasks as part of a new Sci Code benchmark. You will craft original, executable research problems that current frontier models cannot solve and contribute to a benchmark for scientific computing with a focus on chemistry subdomains.
You will source material, write prompts, define grading criteria, and calibrate tasks against models that fail more often than succeed, in a collaborative experiment with leading AI labs.
Mercor is seeking PhD- and Master’s-level scientists to author AI evaluation tasks as part of a new Sci Code benchmark. You will craft original, executable research problems that current frontier models cannot solve and contribute to a benchmark for scientific computing with a focus on chemistry subdomains.
You will source material, write prompts, define grading criteria, and calibrate tasks against models that fail more often than succeed, in a collaborative experiment with leading AI labs.