RL_Notes__final_.pdf
jubayer-ibn-hamid.github.io · 7,030 words · saved by 1 readers
N/A
Deep Reinforcement Learning Jubayer Ibn Hamid 1 Preface This document is a combination of notes I wrote for CS 224R (taught by Prof. Chelsea Finn in Spring, 2025) as the head teaching assistant. The notes contain material taught in CS 234 taught by Prof. Emma Brunskill at Stanford University, CS 285 taught by Prof. Sergey Levine at UC Berkeley, and material from various textbooks and papers (in particular, Richard Sutton and Andrew Barto’s ”Reinforcement Learning: An Introduction” [1]). Although these are reading…
saved by
related reading
- Understanding Policy Gradients | John Lambertjohnwlambert.github.io
- rltheorybook_ABJKS.pdfrltheorybook.github.io
- RLAlgsInMDPs.pdfsites.ualberta.ca
- [2506.22401] Exploration from a Primal-Dual Lens: Value-Incentivized Actor-Critic Methods for Sample-Efficient Online RLarxiv.org
- Policy Gradient Algorithms | Lil'Loglilianweng.github.io
- [1710.10044] Distributional Reinforcement Learning with Quantile Regressionarxiv.org
- [1707.06347] Proximal Policy Optimization Algorithmsarxiv.org
- Deep RL Bootcamp - Lecturessites.google.com
- A (Long) Peek into Reinforcement Learning | Lil'Loglilianweng.github.io
- [1709.06560] Deep Reinforcement Learning that Mattersarxiv.org
- An Updated Introduction to Reinforcement Learning | Sri's Blogsrianumakonda.com
- Part 1: Key Concepts in RL - Spinning Up documentationspinningup.openai.com