Destaca para este puesto — genera un currículum y una carta de presentación adaptados en cuestión de un minuto.
Clera in San Francisco, CA is building core reinforcement learning systems as a founding engineer on an early-stage AI team. You will work directly with the founders and own the process from environment design and model training through evaluation.
You design RL environments, train models with PPO, GRPO, DPO, RLHF, and RLAIF, and develop evaluation frameworks. You will run experiments, interpret results, and decide what approaches to explore next, scaling training on GPU clusters.
Build core reinforcement learning systems from the ground up as a founding engineer on an early-stage AI team. You will work directly with the founders and own the process from environment design and model training through evaluation, helping shape the company's technical direction.
Compensation is $125,000 to $200,000 USD annually.
On-site in San Francisco, California, United States.