1512.04455
arxiv.org · 5,667 words · saved by 1 readers
N/A
Memory-based control with recurrent neural networks Nicolas Heess* Jonathan J Hunt* Timothy P Lillicrap David Silver Google Deepmind arXiv:1512.04455v1 [cs.LG] 14 Dec 2015 * These authors contributed equally. heess,…
saved by
related reading
- Extended Data Fig. 5: Training progress. | Naturenature.com
- 2006.10742arxiv.org
- reinforcement learning - How does one stack multiple observations in the input layer of a convolutional neural network? - Artificial Intelligence Stack Exchangeai.stackexchange.com
- A (Long) Peek into Reinforcement Learning | Lil'Loglilianweng.github.io
- Deep Reinforcement Learning: Pong from Pixelskarpathy.github.io
- Part 1: Key Concepts in RL - Spinning Up documentationspinningup.openai.com
- Deep Q-Networks Explained — LessWronglesswrong.com
- Policy Gradient Algorithms | Lil'Loglilianweng.github.io
- Deep Reinforcement Learning: Pong from Pixelskarpathy.github.io
- Key Papers in Deep RL - Spinning Up documentationspinningup.openai.com
- Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Reviewarxiv.org
- An Updated Introduction to Reinforcement Learning | Sri's Blogsrianumakonda.com