Description
Humanoid is hiring a Reinforcement Learning Engineer for its Autonomy team in London. The role focuses on training language-vision-conditioned manipulation policies through reinforcement learning in simulation and physical reality, developing diverse manipulation tasks, collecting trajectories for behavior cloning, establishing real-world RL training pipelines, and transferring simulation-trained policies to robots. Candidates should have at least three years of deep-learning systems experience, hands-on experience with LLMs, VLMs, or generative models, and strong Python and PyTorch/JAX skills. The position offers equity, paid leave, private healthcare, a pension scheme, and in-office benefits.
