An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Manus AI in Singapore seeks an experienced evaluator to design metrics and experiments for evaluating AI agents and models. You will translate user tasks into measurable criteria, design tasks around planning and tool use, and support post-training comparisons to ensure robust evaluation results.
You will work with product, engineering and model teams to analyze results, identify gaps, and develop new methods that address capability and user-value issues, ensuring rigorous validation.
Manus AI in Singapore seeks an experienced evaluator to design metrics and experiments for evaluating AI agents and models. You will translate user tasks into measurable criteria, design tasks around planning and tool use, and support post-training comparisons to ensure robust evaluation results.
You will work with product, engineering and model teams to analyze results, identify gaps, and develop new methods that address capability and user-value issues, ensuring rigorous validation.