[1611.01779] Learning to Act by Predicting the Future
Abstract:We present an approach to sensorimotor control in immersive environments. Our approach utilizes a high-dimensional sensory stream and a lower-dimensional measurement stream. The cotemporal structure of these streams provides a rich supervisory signal, which enables training a sensorimotor control model by interacting with the environment. The model is trained using supervised learning techniques, but without extraneous supervision. It learns to act based on raw sensory input from a complex three-dimensional environment. The presented formulation enables learning without a fixed goal at training time, and pursuing dynamically changing goals at test time. We conduct extensive experiments in three-dimensional simulations based on the classical first-person game Doom. The results demonstrate that the presented approach outperforms sophisticated prior formulations, particularly on challenging tasks. The results also show that trained models successfully generalize across environments and goals. A model trained using the presented approach won the Full Deathmatch track of the Visual Doom AI Competition, which was held in previously unseen environments.
Published as a conference paper at ICLR 2017 L EARNING TO ACT BY P REDICTING THE F UTURE Alexey Dosovitskiy Vladlen Koltun Intel Labs Intel Labs A BSTRACT We present an approach to sensorimotor control in immersive environments. Our…
saved by
related reading
- pdfopenreview.net
- Dream to Control: Learning Behaviors by Latent Imaginationdanijar.com
- Ch. 21 - Imitation Learningunderactuated.mit.edu
- State of Robot Learning, December 2025vedder.io
- World Models | Rohit Bandarurohitbandaru.github.io
- The flavor of the bitter lesson for computer vision - Vincent Sitzmannvincentsitzmann.com
- LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixelsle-wm.github.io
- How Claude Performs on Robotics Tasks \ Anthropicanthropic.com
- [2509.24527] Training Agents Inside of Scalable World Modelsarxiv.org
- Introducing Dreamer: Scalable Reinforcement Learning Using World Modelsresearch.google
- CIS 6280 · World Modelscis.upenn.edu
- 2006.10742arxiv.org