Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Scale AI, Inc. is recruiting for Research Scientists and Research Engineers specializing in LLM post-training evaluation and benchmark development.
This role focuses on rigorously evaluating frontier models, diagnosing failure modes, and building robust benchmarks for text and multimodal modalities. You will collaborate with researchers and engineers to define best practices in evaluation-driven AI development and translate failure analyses into strategic input for the next generation of
Scale AI, Inc. is recruiting for Research Scientists and Research Engineers specializing in LLM post-training evaluation and benchmark development.
This role focuses on rigorously evaluating frontier models, diagnosing failure modes, and building robust benchmarks for text and multimodal modalities. You will collaborate with researchers and engineers to define best practices in evaluation-driven AI development and translate failure analyses into strategic input for the next generation of