Get more replies from employers
Send a job-specific resume in minutes.
Scale AI is hiring for roles focused on evaluating and benchmarking frontier LLMs and Agents within the GenAI Research Organization. You will develop robust evaluations and RCA-driven insights, collaborating with researchers to shape evaluation-driven AI development and translate failure analyses into strategic input for next-gen models.
The role emphasizes post-training techniques like SFT and RLHF, publication of findings at major AI conferences, and building benchmarks for text and multimodal
Scale AI is hiring for roles focused on evaluating and benchmarking frontier LLMs and Agents within the GenAI Research Organization. You will develop robust evaluations and RCA-driven insights, collaborating with researchers to shape evaluation-driven AI development and translate failure analyses into strategic input for next-gen models.
The role emphasizes post-training techniques like SFT and RLHF, publication of findings at major AI conferences, and building benchmarks for text and multimodal