Stand out for this role — generate a tailored resume and cover letter in about a minute.
Speedrun Talent Network seeks an expert to design, build, and scale distributed RL post-training systems for large-scale models. You will orchestrate trainer, rollout, environment, and reward workloads across thousands of GPUs and integrate inference engines such as vLLM and SGLang.
You will develop RL environments and reward infrastructures, verifiers, and evaluation harnesses, ensuring stable, high-throughput training and rollout across complex, asynchronous pipelines.
Speedrun Talent Network seeks an expert to design, build, and scale distributed RL post-training systems for large-scale models. You will orchestrate trainer, rollout, environment, and reward workloads across thousands of GPUs and integrate inference engines such as vLLM and SGLang.
You will develop RL environments and reward infrastructures, verifiers, and evaluation harnesses, ensuring stable, high-throughput training and rollout across complex, asynchronous pipelines.