Description
Flexion is looking for a Research Engineer specializing in dexterous manipulation and large-scale modeling to lead the development of Vision-Language-Action (VLA) models. This role involves leveraging internet-scale egocentric video to create models enabling humanoid robots to interact with the world like humans. The engineer will architect scalable pre-training objectives, develop multi-modal Foundation Models, design generative policy heads, align human motion with humanoid systems, and utilize offline RL for optimization.

