Get more replies from employers
Send a job-specific resume in minutes.
Mercor collaborates with a leading AI research lab on frontier code agents to assess and improve frontier AI coding models through structured evaluations. Contributors help advance realistic ML engineering workflows and model evaluation, focusing on production-ready deployment and robust inference systems.
The role emphasizes hands-on review of model-generated implementations, bug-hunting, and comparing outputs across frontier models, requiring strong engineering judgment for real-world ML