Enhancing Offline Reinforcement Learning with Curriculum Learning-Based Trajectory Valuation
This is experimental HTML to improve accessibility. We invite you to report rendering errors. Use Alt+Y to toggle on accessible reporting links and Alt+Shift+Y to toggle off. Learn more about this project and help improve conversions. HTML conversions sometimes display errors due to content that did not convert correctly from the source. This paper uses the following packages that are not yet supported by the HTML conversion tool. Feedback on these issues are not necessary; they are known and are being worked on. Authors: achieve the best HTML results from your LaTeX submissions by following these best practices. ifaamas \acmConference[AAMAS ’25]Proc. of the 24th International Conference on Autonomous Agents and Multiagent Systems (AAMAS 2025)May 19 – 23, 2025 Detroit, Michigan, USAY. Vorobeychik, S. Das, A. Nowé (eds.) \copyrightyear2025 \acmYear2025 \acmDOI \acmPrice \acmISBN \acmSubmissionID6 \affiliation \institutionL3S Research Center \cityHannover \countryGermany \affiliation \i
\setcopyright ifaamas \acmConference [AAMAS ’25]Proc. of the 24th International Conference on Autonomous Agents and Multiagent Systems (AAMAS 2025)May 19 – 23, 2025 Detroit, Michigan, USAY. Vorobeychik, S. Das, A. Nowé (eds.) \copyrightyear 2025 \acmYear 2025 \acmDOI \acmPrice \acmISBN \acmSubmissionID 6 \affiliation \institution L3S Research Center \city Hannover \country Germany \affiliation \institution Technical University of Berlin \city Berlin \country Germany \affiliation \institution Delft University of Technology \city Delft \country Netherlands \affiliation \institution L3S Research
Explore this link on the map →related reading
- [2005.01643] Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problemsar5iv.labs.arxiv.org
- [2204.05618] When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?ar5iv.labs.arxiv.org
- On-Policy Distillation - Thinking Machines Labthinkingmachines.ai
- A (Long) Peek into Reinforcement Learning | Lil'Loglilianweng.github.io
- [2406.09329] Is Value Learning Really the Main Bottleneck in Offline RL?ar5iv.labs.arxiv.org
- [2201.11861] The Challenges of Exploration for Offline Reinforcement Learningar5iv.labs.arxiv.org
- Should I Use Offline RL or Imitation Learning? – The Berkeley Artificial Intelligence Research Blogbair.berkeley.edu
- Exploration for the Efficient Deployment of Reinforcement Learning Agentsopenreview.net
- [2006.03647] Deployment-Efficient Reinforcement Learning via Model-Based Offline Optimizationar5iv.labs.arxiv.org
- pistar06.pdfpi.website
- [2307.04354] Policy Finetuning in Reinforcement Learning via Design of Experiments using Offline Dataar5iv.labs.arxiv.org
- An Updated Introduction to Reinforcement Learning | Sri's Blogsrianumakonda.com