Verschicke keinen generischen Lebenslauf — erstelle einen Lebenslauf und ein Anschreiben, die genau auf diese Rolle zugeschnitten sind.
Mercor is hiring PhD and Master's scientists to author AI evaluation tasks for Sci Code, collaborating with leading AI labs on a new scientific benchmark. You will craft original, executable research problems that current frontier models cannot solve, emphasizing depth in multiple physics subdomains and coding for science.
You will source material, write prompts, and design grading criteria; you will calibrate with frontier models to ensure robust failure modes.
Mercor is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code)
Mercor is partnering with leading AI labs on a new benchmark for scientific computing. You will author original, executable research problems that today's frontier models cannot solve.
Source your own material: a published paper, a Kaggle dataset, an open-source repository, or a scenario you design
Write scientific prompts based on the input
Build the grading criteria that define a correct answer
Calibrate against frontier models — a task ships only when strong models fail it more often than they succeed
PhD in physics, applied physics, or a closely related field
Demonstrated depth in at least two of the following subdomains: condensed matter, optics, quantum information/computing, computational physics, astrophysics, particle physics
Working proficiency in Python, R, or another relevant programming language for scientific computing
Comfortable with Git/GitHub and running code in Docker — authoring runs through a pull-request workflow with automated quality checks
Publications in peer-reviewed journals
Prior scientific software or research engineering experience
Duration: 6 weeks
Commitment: part-time, 20+ hours per week
Start date: immediate
Upload your resume and application form
A 25-minute conversational interview covering your background, experience, and motivations
Follow up within a few days with next steps and onboarding