✳flâneur — a map of the web's best reading
Reward Isn’t Free: Supervising Robot Learning with Language and Video from the Web | SAIL Blog
ai.stanford.edu · 2,363 words · saved by 1 readers
This work was conducted as part of SAIL and CRFM.
This work was conducted as part of SAIL and CRFM . Deep learning has enabled improvements in the capabilities of robots on a range of problems such as grasping 1 and locomotion 2 in recent years. However, building the quintessential home robot that can perform a range of interactive tasks, from cooking to cleaning, in novel environments has remained elusive. While a number of hardware and software challenges remain, a necessary component is robots that can generalize their prior knowledge to new environments, tasks, and objects in a zero or few shot manner. For example, a home robot tasked wit
Explore this link on the map →saved by
related reading
- [2109.01115] Learning Language-Conditioned Robot Behavior from Offline Data and Crowd-Sourced Annotationarxiv.org
- Emergence of Human to Robot Transfer in Vision-Language-Action Modelspi.website
- State of Robot Learning, December 2025vedder.io
- Just Ask for Generalization | Eric Jangevjang.com
- how we accidentally solved robotics by watching 1 million hours of YouTube – atharva's blogksagar.bearblog.dev
- Generalist - GEN-0 / Embodied Foundation Models That Scale with Physical Interactiongeneralistai.com
- Causal Video Models Are Data-Efficient Robot Policy Learners | Rhoda AIrhoda.ai
- A VLA with Open-World Generalizationpi.website
- To Understand Language is to Understand Generalization | Eric Jangevjang.com
- A Steerable Model with Emergent Capabilitiespi.website
- Open X-Embodiment: Robotic Learning Datasets and RT-X Modelsarxiv.org
- [2206.11795] Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videosarxiv.org