An application made for this job — a tailored resume and cover letter that speak straight to the posting.
NVIDIA is seeking an RL post-training infrastructure engineer to design and scale rollout loops for AI research and production workloads. You will optimize RL training, inference, and rollout across GPU, CPU, and LPU resources while improving open-source frameworks and runtimes.
You will collaborate with researchers, labs, and hardware teams to drive resilient, scalable systems and contribute to VeRL, Miles, TorchTitan, Ray, and Monarch in a frontier AI environment.
NVIDIA is seeking an RL post-training infrastructure engineer to design and scale rollout loops for AI research and production workloads. You will optimize RL training, inference, and rollout across GPU, CPU, and LPU resources while improving open-source frameworks and runtimes.
You will collaborate with researchers, labs, and hardware teams to drive resilient, scalable systems and contribute to VeRL, Miles, TorchTitan, Ray, and Monarch in a frontier AI environment.