Understanding the role of the discount factor in reinforcement learning - Cross Validated
I'm teaching myself about reinforcement learning, and trying to understand the concept of discounted reward. So the reward is necessary to tell the system which state-action pairs are good, and whi...
Understanding the role of the discount factor in reinforcement learning - Cross Validated Stack Internal Knowledge at work Bring the best of human thought and AI automation together at your work. Explore Stack Internal Understanding the role of the discount factor in reinforcement learning Ask Question Asked 9 years, 11 months ago Modified 4 years, 11 months ago Viewed 145k times 99 $\begingroup$ I'm teaching myself about reinforcement learning, and trying to understand the concept of discounted reward. So the reward is necessary to tell the system which state-action pairs are good, and which
saved by
related reading
- Reward is not the optimization target — LessWronglesswrong.com
- Part 1: Key Concepts in RL - Spinning Up documentationspinningup.openai.com
- A (Long) Peek into Reinforcement Learning | Lil'Loglilianweng.github.io
- Temporal discounting - EA Forumforum.effectivealtruism.org
- RUDDER - Reinforcement Learning with Delayed Rewards | rudderml-jku.github.io
- A Crash Course in the Neuroscience of Human Motivation — LessWronglesswrong.com
- Reward Is Not the Optimization Targetturntrout.com
- Models Don't "Get Reward" — LessWronglesswrong.com
- An Updated Introduction to Reinforcement Learning | Sri's Blogsrianumakonda.com
- RL_Notes__final_.pdfjubayer-ibn-hamid.github.io
- Hyperbolic discounting - The Decision Labthedecisionlab.com
- Reinforcement learning - Wikipediaen.wikipedia.org