Get more replies from employers
Send a job-specific resume in minutes.
Innodata Inc. is seeking Reward Validation Specialists to contribute to advanced AI training and evaluation projects. PhD-qualified researchers with RL, optimization, and Python skills will help improve the reliability and accuracy of AI evaluation systems.
You will build and validate automated grading systems for AI model evaluation pipelines, focusing on step-by-step reward logic and scalable Python tools to ensure consistent, objective assessments across large workloads.
We are hiring Reward Validation Specialists to contribute to advanced AI training and evaluation projects. This role is ideal for PhD-qualified researchers with expertise in reinforcement learning, optimization, and Python programming who are passionate about improving the reliability and accuracy of AI evaluation systems.
In this role, you will help build and validate automated grading systems for AI model evaluation pipelines. Rather than assessing only the final output, you will evaluate and verify step-by-step reward logic to ensure each grading component accurately measures the intended behavior. You will also develop and implement automated graders in Python that can scale across large evaluation workflows while maintaining consistency, accuracy, and alignment with task requirements.