Get more replies from employers
Send a job-specific resume in minutes.
Prime Recruitment Partners is partnering with a well-funded, early-stage AI lab to advance fundamental research at the intersection of reinforcement learning, reasoning, and computational problems.
You will build and train RL agents, design curricula and reward structures, and develop evaluation frameworks to separate genuine reasoning from mere benchmark optimization. The role scales training across distributed GPU infrastructure and collaborates with world-class researchers.
We're partnering with a well-funded, early-stage AI lab building research systems aimed at one of the hardest open problems in the field: getting models to do real discovery work rather than pattern‑match their way through benchmarks. The team is drawn from leading AI labs and top research institutions. This is not product engineering or benchmark chasing. It's fundamental research at the intersection of reinforcement learning, reasoning, and computational research problems, in an environment built around exploration, publication, and close collaboration with world‑class researchers and engineers.