Destaca-te para esta função — gera um currículo e uma carta de apresentação personalizados em cerca de um minuto.
Datamentors seeks an Reinforcement Learning Lead to own the post-training stage of the Ardia humanoid platform, lifting success, precision and robustness through RL fine-tuning, reward modelling and a tight sim-to-real loop. You will set the technical direction, build training/evaluation infra, and grow a small team around it.
You will lead the RL roadmap, mentor engineers, and collaborate with perception and control teams to turn demonstrations into a dependable product.
Datamentors seeks an Reinforcement Learning Lead to own the post-training stage of the Ardia humanoid platform, lifting success, precision and robustness through RL fine-tuning, reward modelling and a tight sim-to-real loop. You will set the technical direction, build training/evaluation infra, and grow a small team around it.
You will lead the RL roadmap, mentor engineers, and collaborate with perception and control teams to turn demonstrations into a dependable product.