Helix: A Vision-Language-Action Model for Generalist Humanoid Control
figure.ai · 2,159 words · saved by 4 readers
Figure was founded with the ambition to change the world.
Helix: A Vision-Language-Action Model for Generalist Humanoid Control February 20, 2025 Introducing Helix We're introducing Helix, a generalist Vision-Language-Action (VLA) model that unifies perception, language understanding, and learned control to overcome multiple longstanding challenges in robotics. Helix is a series of firsts: Full-upper-body control : Helix is the first VLA to output high-rate continuous control of the entire humanoid upper body, including wrists, torso, head, and individual fingers. Multi-robot collaboration : Helix is the first VLA to operate simultaneously on two rob
saved by
related reading
- How Claude Performs on Robotics Tasks \ Anthropicanthropic.com
- GitHub - robotics-survey/Awesome-Robotics-Foundation-Modelsgithub.com
- LeRobot v0.6.0: Imagine, Evaluate, Improvehuggingface.co
- A VLA with Open-World Generalizationpi.website
- Emergence of Human to Robot Transfer in Vision-Language-Action Modelspi.website
- Show-Harnessshowlab.github.io
- RT-2: Vision-Language-Action Modelsrobotics-transformer2.github.io
- $π_0$: A Vision-Language-Action Flow Model for General Robot Controlalphaxiv.org
- Vision-Language-Action (VLA) Models: A Review of Recent Progressxxxxyu.github.io
- A Steerable Model with Emergent Capabilitiespi.website
- State of Vision-Language-Action (VLA) Research at ICLR 2026 – Moritz Reussmbreuss.github.io
- Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Controlsteerable-policies.github.io