Policy Finetuning in Reinforcement Learning via Design of Experiments using Offline Data | HTML5
In some applications of reinforcement learning, a dataset of pre-collected experience is already available but it is also possible to acquire some additional online data to help improve the quality of the policy. Howev…
[2307.04354] Policy Finetuning in Reinforcement Learning via Design of Experiments using Offline Data Policy Finetuning in Reinforcement Learning via Design of Experiments using Offline Data Ruiqi Zhang Department of Statistics University of California Berkeley rqzhang@berkeley.edu Andrea Zanette Department of EECS University of California Berkeley zanette@berkeley.edu Abstract In some applications of reinforcement learning, a dataset of pre-collected experience is already available but it is also possible to acquire some additional online data to help improve the quality of the policy. Howeve
Explore this link on the map →related reading
- Exploration for the Efficient Deployment of Reinforcement Learning Agentsopenreview.net
- A (Long) Peek into Reinforcement Learning | Lil'Loglilianweng.github.io
- [2005.01643] Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problemsar5iv.labs.arxiv.org
- [2201.11861] The Challenges of Exploration for Offline Reinforcement Learningar5iv.labs.arxiv.org
- Part 1: Key Concepts in RL - Spinning Up documentationspinningup.openai.com
- [2007.08202] Provably Good Batch Reinforcement Learning Without Great Explorationar5iv.labs.arxiv.org
- [2204.05618] When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?ar5iv.labs.arxiv.org
- [2006.03647] Deployment-Efficient Reinforcement Learning via Model-Based Offline Optimizationar5iv.labs.arxiv.org
- Reinforcement Learning via Implicit Imitation Guidancearxiv.org
- [2406.09329] Is Value Learning Really the Main Bottleneck in Offline RL?ar5iv.labs.arxiv.org
- Is Offline Decision Making Possible with Only Few Samples? Reliable Decisions in Data-Starved Bandits via Trust Region Enhancementarxiv.org
- An Updated Introduction to Reinforcement Learning | Sri's Blogsrianumakonda.com