Get more replies from employers
Send a job-specific resume in minutes.
Mercor in San Francisco seeks an experienced software development evaluator to assess AI-assisted coding traces used to train and evaluate frontier AI models. You will review end-to-end coding sessions produced with AI-assisted developer tools, judging correctness, workflow soundness, and reasoning, and deliver clear rubric-based feedback.
The role emphasizes collaboration with ML researchers and software engineers; you will apply structured criteria, document findings, and help improve tooling
Evaluate the quality and correctness of AI-assisted software-development traces used to train and evaluate a frontier AI lab's models. You'll assess end-to-end coding sessions produced with AI-assisted developer tools — judging correctness, workflow soundness, and reasoning — and provide clear, rubric-based written feedback.