Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Control
steerable-policies.github.io · 1,118 words · saved by 1 readers
Project page for Steerable Vision-Language-Action Policies
Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Control Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Control William Chen 1 , Jagdeep Bhatia 1 , Catherine Glossop 1 , Nikhil Mathihalli 1 , Ria Doshi 2 , Andy Tang 2 , Danny Driess 3 , Karl Pertsch 3 , Sergey Levine 1 1 UC Berkeley, 2 Stanford University 3 Physical Intelligence Paper Code Model Data Preview (Muted) Full (Narrated) Steerable Policies can flexibly follow diverse commands, allowing them to better interface with VLMs to transfer foundation model capabilities to t
saved by
related reading
- 1X World Model | From Video to Action: A New Way Robots Learn1x.tech
- SayCan: Grounding Language in Robotic Affordancessay-can.github.io
- Causal Video Models Are Data-Efficient Robot Policy Learners | Rhoda AIrhoda.ai
- How Claude Performs on Robotics Tasks \ Anthropicanthropic.com
- A Steerable Model with Emergent Capabilitiespi.website
- Helix: A Vision-Language-Action Model for Generalist Humanoid Controlfigure.ai
- Pretrained to Imagine, Fine-Tuned to Act: The Rise of World-Action Models | NVIDIA Technical Blogdeveloper.nvidia.com
- State of Vision-Language-Action (VLA) Research at ICLR 2026 – Moritz Reussmbreuss.github.io
- RT-2: Vision-Language-Action Modelsrobotics-transformer2.github.io
- A VLA with Open-World Generalizationpi.website
- Vision-Language-Action (VLA) Models: A Review of Recent Progressxxxxyu.github.io
- $π_0$: A Vision-Language-Action Flow Model for General Robot Controlalphaxiv.org