✳flâneur — a map of the web's best reading
SARM2: Multi-Task Stage Aware Reward Modeling for Self Improving Robotic Manipulation
qianzhong-chen.github.io · 757 words · saved by 1 readers
SARM2: Multi-Task Stage Aware Reward Modeling for Self Improving Robotic Manipulation
SARM2: Multi-Task Stage Aware Reward Modeling for Self Improving Robotic Manipulation --> --> SARM2: Multi-Task Stage Aware Reward Modeling for Self Improving Robotic Manipulation Qianzhong Chen 1 , Hau Zheng 1 , Justin Yu 2 , Suning Huang 1 , Jiankai Sun 1 , Ken Goldberg 2 , Pieter Abbeel 2 , Chuan Wen 3 , Yide Shentu 2, 4 , Philipp Wu 4 Mac Schwager 1 , 1 Stanford University, 2 UC Berkeley, 3 Shanghai Jiao Tong University, 4 xdof.ai Accepted to ICLR 2026 --> Try SARM, now native to LeRobot ! Thanks to Hugging Face team for the support! --> Paper --> SARM (prev. Work) arXiv twitter Code (comi
Explore this link on the map →saved by
related reading
- Precise Manipulation with Efficient Online RLpi.website
- Explore | alphaXivalphaxiv.org
- pistar06.pdfpi.website
- State of Robot Learning, December 2025vedder.io
- e5b5c402bb7bd5e60bede6961d6fe39e-Paper-Conference.pdfproceedings.iclr.cc
- A VLA with Open-World Generalizationpi.website
- RT-2: Vision-Language-Action Modelsrobotics-transformer2.github.io
- π*0.6: a VLA That Learns From Experiencephysicalintelligence.company
- A Steerable Model with Emergent Capabilitiespi.website
- Causal Video Models Are Data-Efficient Robot Policy Learners | Rhoda AIrhoda.ai
- MolmoAct Action Reasoning Models that can Reason in Spacearxiv.org
- A VLA that Learns from Experiencepi.website