Get more replies from employers
Send a job-specific resume in minutes.
1X in San Carlos, California, is looking for an expert in reinforcement learning to develop and deploy policies for NEO, a humanoid robot. You will be responsible for closing the gap between simulation and real-world performance, ensuring reliable operation in home environments.
The ideal candidate has strong skills in Python or C++, experience with PyTorch, and a background in training RL policies. This role offers a competitive compensation package with comprehensive benefits including medical coverage and generous PTO.
We’re building humanoid robots that work in home - doing the chores, handling the tasks, and giving people their time back. Simple, but it’s not.
To do this right, we have to solve robotics, AI, manufacturing - at the same time, at scale, in a form factor that has to be safe enough to live with your family. If you’re inspired by this, you’ll thrive here. We’ve been at this since 2014 and we’re at the point where the hard problems are behind us and the hard work is in front of us.
NEO is our flagship - a home robot designed to move, learn, and operate in the real world alongside real people. We’re not demoing it - we’re shipping it. We’re excited to meet you, if this excites you.
If you’ve spent your career working on problems that matter and want to see them actually reach the world - this is that moment. We’re scaling, we’re hiring with intention, and we need people who want to build something that will genuinely change how humans spend their time - safely creating abundance for all.
The Reinforcement Learning team teaches NEO new capabilities, training policies for manipulation and locomotion tasks across simulation and real-world environments, then deploying them into homes. We work at the intersection of algorithm development, sim-to-real transfer, and production deployment: our research is only successful when a policy runs reliably on a physical robot in the field. If you want to directly expand what a humanoid robot can do for people, this is that team.
Own the full pipeline from RL algorithm development through production deployment—training NEO on manipulation and locomotion tasks in simulation, closing the sim-to-real gap, and shipping policies that work reliably in real-world home environments. This is critical-path work: the range of tasks NEO can perform safely and reliably is a direct function of the quality of RL policies your team ships. You will collaborate closely with hardware, controls, data collection, and QA teams, and measure your impact by what NEO can do in the field.
$200,000 - $300,000 + Equity
1X is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, gender, gender identity or expression, sexual orientation, national origin, ancestry, citizenship, age, marital status, medical condition, genetic information, disability, military or veteran status, justice system impact, or any other characteristic protected under applicable federal, state, or local law.