Get more replies from employers
Send a job-specific resume in minutes.
Mercor is seeking PhD and Master’s scientists to author AI evaluation tasks for a new benchmark in scientific computing. You will source material, craft executable problems, and build the grading criteria to measure model performance on challenging scenarios.
You will calibrate prompts against frontier models, ensuring tasks remain difficult where state-of-the-art systems struggle, across two subdomains with a coding focus.
Mercor is seeking PhD and Master’s scientists to author AI evaluation tasks for a new benchmark in scientific computing. You will source material, craft executable problems, and build the grading criteria to measure model performance on challenging scenarios.
You will calibrate prompts against frontier models, ensuring tasks remain difficult where state-of-the-art systems struggle, across two subdomains with a coding focus.