Deep Reinforcement Learning in a Handful of Trials using Probabilistic Dynamics Models | HTML5
Model-based reinforcement learning (RL) algorithms can attain excellent sample efficiency, but often lag behind the best model-free algorithms in terms of asymptotic performance. This is especially true with high-capac…
Deep Reinforcement Learning in a Handful of Trials using Probabilistic Dynamics Models Kurtland Chua Roberto Calandra Rowan McAllister Sergey Levine Berkeley Artificial Intelligence Research University of California, Berkeley {kchua, roberto.calandra, rmcallister, svlevine}@berkeley.edu Abstract Model-based reinforcement learning (RL) algorithms can attain excellent sample efficiency, but often lag behind the best model-free algorithms in terms of asymptotic performance. This is especially true with high-capacity parametric function approximators, such as deep networks. In this paper, we study
related reading
- Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Reviewarxiv.org
- Part 2: Kinds of RL Algorithms - Spinning Up documentationspinningup.openai.com
- A (Long) Peek into Reinforcement Learning | Lil'Loglilianweng.github.io
- Key Papers in Deep RL - Spinning Up documentationspinningup.openai.com
- [2006.03647] Deployment-Efficient Reinforcement Learning via Model-Based Offline Optimizationar5iv.labs.arxiv.org
- Model-free (reinforcement learning) - Wikipediaen.wikipedia.org
- [1709.06560] Deep Reinforcement Learning that Mattersarxiv.org
- NeurIPS-2021-understanding-end-to-end-model-based-reinforcement-learning-methods-as-implicit-parameterization-Supplemental.pdflis.csail.mit.edu
- [1802.09081] Temporal Difference Models: Model-Free Deep RL for Model-Based Controlar5iv.labs.arxiv.org
- State of RL for reasoning LLMs | A. Weersaweers.de
- c21f4ce780c5c9d774f79841b81fdc6d-Paper.pdfproceedings.neurips.cc
- Deep RL Bootcamp - Lecturessites.google.com