Learning Generalizable Feature Fields for Mobile Manipulation
This is experimental HTML to improve accessibility. We invite you to report rendering errors. Use Alt+Y to toggle on accessible reporting links and Alt+Shift+Y to toggle off. Learn more about this project and help improve conversions.
\addbibresource main.bib Learning Generalizable Feature Fields for Mobile Manipulation Ri-Zhao Qiu ∗1 , Yafei Hu ∗1,2 , Yuchen Song ∗1 , Ge Yang 3 , Yang Fu 1 , Jianglong Ye 1 , Jiteng Mu 1 , Ruihan Yang 1 , Nikolay Atanasov 1 , Sebastian Scherer 2 , Xiaolong Wang 1 ∗ equal contribution 1 UC San Diego 2 CMU 3 MIT https://geff-b1.github.io Abstract An open problem in mobile manipulation is how to represent objects and scenes in a unified manner so that robots can use both for navigation and manipulation. The latter requires capturing intricate geometry while understanding fine-grained semantics
Explore this link on the map →related reading
- NeRF: Neural Radiance Fieldsmatthewtancik.com
- When Models Manipulate Manifolds: The Geometry of a Counting Tasktransformer-circuits.pub
- The flavor of the bitter lesson for computer vision - Vincent Sitzmannvincentsitzmann.com
- Feature Visualizationdistill.pub
- A VLA with Open-World Generalizationpi.website
- ReferIt3D: Neural Listeners for Fine-Grained 3D Object Identification in Real-World Scenesecva.net
- Machine Learning for Inverse Graphics – Scene Representation Groupscenerepresentations.org
- A Steerable Model with Emergent Capabilitiespi.website
- Feature-wise transformationsdistill.pub
- e5b5c402bb7bd5e60bede6961d6fe39e-Paper-Conference.pdfproceedings.iclr.cc
- PLA: Language-Driven Open-Vocabulary 3D Scene Understandingarxiv.org
- Sparse Autoencoders Reveal Interpretable and Steerable Features in VLA Modelsarxiv.org