SARM2: Multi-Task Stage Aware Reward Modeling for Self Improving Robotic Manipulation
qianzhong-chen.github.io · 757 words · saved by 1 readers
SARM2: Multi-Task Stage Aware Reward Modeling for Self Improving Robotic Manipulation
SARM2: Multi-Task Stage Aware Reward Modeling for Self Improving Robotic Manipulation --> --> SARM2: Multi-Task Stage Aware Reward Modeling for Self Improving Robotic Manipulation Qianzhong Chen 1 , Hau Zheng 1 , Justin Yu 2 , Suning Huang 1 , Jiankai Sun 1 , Ken Goldberg 2 , Pieter Abbeel 2 , Chuan Wen 3 , Yide Shentu 2, 4 , Philipp Wu 4 Mac Schwager 1 , 1 Stanford University, 2 UC Berkeley, 3 Shanghai Jiao Tong University, 4 xdof.ai Accepted to ICLR 2026 --> Try SARM, now native to LeRobot ! Thanks to Hugging Face team for the support! --> Paper --> SARM (prev. Work) arXiv twitter Code (comi
saved by
related reading
- How Claude Performs on Robotics Tasks \ Anthropicanthropic.com
- Precise Manipulation with Efficient Online RLpi.website
- SimpleVLA-RL: Scaling VLA Training via Reinforcement Learningalphaxiv.org
- Explore | alphaXivalphaxiv.org
- pistar06.pdfpi.website
- 2310.12921.pdfarxiv.org
- State of Robot Learning, December 2025vedder.io
- A Steerable Model with Emergent Capabilitiespi.website
- e5b5c402bb7bd5e60bede6961d6fe39e-Paper-Conference.pdfproceedings.iclr.cc
- Pretrained to Imagine, Fine-Tuned to Act: The Rise of World-Action Models | NVIDIA Technical Blogdeveloper.nvidia.com
- A VLA with Open-World Generalizationpi.website
- Helix: A Vision-Language-Action Model for Generalist Humanoid Controlfigure.ai