𝜋₀: A Vision-Language-Action Flow Model for General Robot Control
This is experimental HTML to improve accessibility. We invite you to report rendering errors. Use Alt+Y to toggle on accessible reporting links and Alt+Shift+Y to toggle off. Learn more about this project and help improve conversions.
π 0 subscript 𝜋 0 \pi_{0} italic_π start_POSTSUBSCRIPT 0 end_POSTSUBSCRIPT : A Vision-Language-Action Flow Model for General Robot Control Physical Intelligence Kevin Black, Noah Brown, Danny Driess, Adnan Esmail, Michael Equi, Chelsea Finn, Niccolo Fusai, Lachy Groom, Karol Hausman, Brian Ichter, Szymon Jakubczak, Tim Jones, Liyiming Ke, Sergey Levine, Adrian Li-Bell, Mohith Mothukuri, Suraj Nair, Karl Pertsch, Lucy Xiaoyang Shi, James Tanner, Quan Vuong, Anna Walling, Haohuan Wang, Ury Zhilinsky https://physicalintelligence.company/blog/pi0 Physical Intelligence, San Francisco, California,
related reading
- $π_0$: A Vision-Language-Action Flow Model for General Robot Controlalphaxiv.org
- A Steerable Model with Emergent Capabilitiespi.website
- A VLA with Open-World Generalizationpi.website
- Emergence of Human to Robot Transfer in Vision-Language-Action Modelspi.website
- Pretrained to Imagine, Fine-Tuned to Act: The Rise of World-Action Models | NVIDIA Technical Blogdeveloper.nvidia.com
- How Claude Performs on Robotics Tasks \ Anthropicanthropic.com
- Generalist - GEN-0 / Embodied Foundation Models That Scale with Physical Interactiongeneralistai.com
- 45d74e190008c7bff2845ffc8e3facd3-Paper-Conference.pdfproceedings.iclr.cc
- State of Vision-Language-Action (VLA) Research at ICLR 2026 – Moritz Reussmbreuss.github.io
- Physical Intelligence (π)pi.website
- RT-2: Vision-Language-Action Modelsrobotics-transformer2.github.io
- [2602.10556] LAP: Language-Action Pre-Training Enables Zero-shot Cross-Embodiment Transferarxiv.org