An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Inception seeks experienced scientists and engineers with deep expertise in post-training large language models through reinforcement learning. You will design and implement RL training pipelines for diffusion LLMs, develop reward modeling strategies, and build the algorithms that align model behavior with human intent at scale.
The role focuses on designing RL training pipelines, building reward models, and advancing alignment techniques for enterprise-grade LLMs, with emphasis on stability,
We seek experienced scientists and engineers with deep expertise in post-training large language models through reinforcement learning. You will design and implement RL training pipelines for our diffusion LLMs, develop reward modeling strategies, and build the algorithms that align model behavior with human intent at scale.