[2006.10701] Deep Reinforcement Learning amidst Lifelong Non-Stationarity
As humans, our goals and our environment are persistently changing throughout our lifetime based on our experiences, actions, and internal and external drives. In contrast, typical reinforcement learning problem set-up…
Deep Reinforcement Learning amidst Lifelong Non-Stationarity Annie Xie, James Harrison, Chelsea Finn Stanford University, Stanford, CA {anniexie,jharrison,cbfinn}@stanford.edu Abstract As humans, our goals and our environment are persistently changing throughout our lifetime based on our experiences, actions, and internal and external drives. In contrast, typical reinforcement learning problem set-ups consider decision processes that are stationary across episodes. Can we develop reinforcement learning algorithms that can cope with the persistent change in the former, more realistic problem se
Explore this link on the map →related reading
- A (Long) Peek into Reinforcement Learning | Lil'Loglilianweng.github.io
- Part 1: Key Concepts in RL - Spinning Up documentationspinningup.openai.com
- Key Papers in Deep RL - Spinning Up documentationspinningup.openai.com
- Reinforcement learning - Wikipediaen.wikipedia.org
- State of RL for reasoning LLMs | A. Weersaweers.de
- pdfopenreview.net
- An Updated Introduction to Reinforcement Learning | Sri's Blogsrianumakonda.com
- pistar06.pdfpi.website
- Training agents to plan in latent space — a technical overview | by Lukas Bierling | Mediummedium.com
- [1805.12114] Deep Reinforcement Learning in a Handful of Trials using Probabilistic Dynamics Modelsar5iv.labs.arxiv.org
- [1707.06347] Proximal Policy Optimization Algorithmsarxiv.org
- [2005.01643] Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problemsar5iv.labs.arxiv.org