Robotic Task Ambiguity Resolution via Natural Language Interaction
This is experimental HTML to improve accessibility. We invite you to report rendering errors. Use Alt+Y to toggle on accessible reporting links and Alt+Shift+Y to toggle off. Learn more about this project and help improve conversions.
Robotic Task Ambiguity Resolution via Natural Language Interaction Eugenio Chisari 1 , Jan Ole von Hartz 1 , Fabien Despinoy 2 , and Abhinav Valada 1 1 Department of Computer Science, University of Freiburg, Germany. 2 Toyota Motor Europe.This work was funded by Toyota Motor Europe. Abstract Language-conditioned robotic policies allow users to specify tasks using natural language. While much research has focused on improving the action prediction of language-conditioned policies, reasoning about task descriptions has been largely overlooked. Ambiguous task descriptions often lead to downstream
Explore this link on the map →related reading
- Language Models can Solve Computer Tasksarxiv.org
- A VLA with Open-World Generalizationpi.website
- A Steerable Model with Emergent Capabilitiespi.website
- RT-2: Vision-Language-Action Modelsrobotics-transformer2.github.io
- Explore | alphaXivalphaxiv.org
- how we accidentally solved robotics by watching 1 million hours of YouTube – atharva's blogksagar.bearblog.dev
- SayCan: Grounding Language in Robotic Affordancessay-can.github.io
- Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Controlsteerable-policies.github.io
- Helix: A Vision-Language-Action Model for Generalist Humanoid Controlfigure.ai
- [2109.01115] Learning Language-Conditioned Robot Behavior from Offline Data and Crowd-Sourced Annotationarxiv.org
- Precise Manipulation with Efficient Online RLpi.website
- State of Vision-Language-Action (VLA) Research at ICLR 2026 – Moritz Reussmbreuss.github.io