Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
HumanitApp is seeking an evaluator to review the quality and rigor of applied ML tasks used to train and assess frontier AI models. You will examine experiment design, model-selection reasoning, and evaluation methodology, delivering clear rubric-based feedback to support research teams.
The role requires 3+ years in AI evaluation or related fields, with emphasis on rigorous methodological critique and precise written communication. Expect a 40-hour work week at a competitive hourly rate.
Evaluate the quality, correctness, and methodological rigor of applied machine-learning tasks used to train and evaluate a frontier AI lab's models. You'll assess experiment design, model-selection reasoning, and evaluation methodology
and provide clear, rubric-based written feedback.
Basic Qualifications • 3+ y...
AI evaluation Machine learning Research
This opportunity may suit professionals with relevant experience in AI evaluation, Machine learning, Research. Review the official description and requirements before applying.
The listing states $70 - $90 / hour. Confirm the final rate, workload, and payment terms during the official application process.
The listing states $70 - $90 / hour. Confirm the final rate, workload, and payment terms during the official application process.