flâneur — a map of the web's best reading

RT-2: Vision-Language-Action Models

robotics-transformer2.github.io · 1,592 words · saved by 2 readers

Project page for RT-2

RT-2: Vision-Language-Action Models --> --> Your browser does not support the video tag. RT2: Vision-Language-Action Models RT-2 model picking up object given the prompt "pick up the extinct animal." RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control Anthony Brohan Noah Brown Justice Carbajal Yevgen Chebotar Xi Chen Krzysztof Choromanski Tianli Ding Danny Driess Avinava Dubey Chelsea Finn Pete Florence Chuyuan Fu Montse Gonzalez Arenas Keerthana Gopalakrishnan Kehang Han Karol Hausman Alex Herzog Jasmine Hsu Brian Ichter Alex Irpan Nikhil Joshi Ryan Julian Dmitry Kal

Explore this link on the map →

saved by

related reading