✳flâneur — a map of the web's best reading
c21f4ce780c5c9d774f79841b81fdc6d-Paper.pdf
proceedings.neurips.cc · 8,033 words · saved by 1 readers
N/A
Sample-Efficient Reinforcement Learning for Linearly-Parameterized MDPs with a Generative Model Bingyan Wang∗ Yuling Yan∗ Princeton University Princeton University bingyanw@princeton.edu yulingy@princeton.edu Jianqing Fan Princeton University jqfan@princeton.edu…
Explore this link on the map →saved by
related reading
- rltheorybook_ABJKS.pdfrltheorybook.github.io
- A (Long) Peek into Reinforcement Learning | Lil'Loglilianweng.github.io
- Q-learning is not yet scalableseohong.me
- Part 2: Kinds of RL Algorithms - Spinning Up documentationspinningup.openai.com
- RLAlgsInMDPs.pdfsites.ualberta.ca
- A Free Lunch from the Noise:Provable and Practical Exploration for Representation Learningarxiv.org
- 2203.00543arxiv.org
- [1802.09081] Temporal Difference Models: Model-Free Deep RL for Model-Based Controlar5iv.labs.arxiv.org
- Reinforcement Learning in Newcomblike Problemsproceedings.neurips.cc
- Key Papers in Deep RL - Spinning Up documentationspinningup.openai.com
- Reinforcement learning - Wikipediaen.wikipedia.org
- Q-learning - Wikipediaen.wikipedia.org