Reinforcement learning
Reinforcement learning (RL) is an area of machine learning concerned with how intelligent agents ought to take actions in an environment in order to maximize the notion of cumulative reward. Reinforcement learning is one of three basic machine learning paradigms, alongside supervised learning and unsupervised learning.
Reinforcement learning - Wikipedia Jump to content From Wikipedia, the free encyclopedia Field of machine learning For reinforcement learning in psychology, see Reinforcement and Operant conditioning . The typical framing of a reinforcement learning (RL) scenario: an agent takes actions in an environment, which is interpreted into a reward and a state representation, which are fed back to the agent. Part of a series on Machine learning and data mining Paradigms Supervised learning Unsupervised learning Semi-supervised learning Self-supervised learning Reinforcement learning Meta-learning Onlin
Explore this link on the map →related reading
- A (Long) Peek into Reinforcement Learning | Lil'Loglilianweng.github.io
- Part 1: Key Concepts in RL - Spinning Up documentationspinningup.openai.com
- An Updated Introduction to Reinforcement Learning | Sri's Blogsrianumakonda.com
- Reward is not the optimization target — LessWronglesswrong.com
- Key Papers in Deep RL - Spinning Up documentationspinningup.openai.com
- Part 1: Key Concepts in RL - Spinning Up documentationspinningup.openai.com
- State of RL for reasoning LLMs | A. Weersaweers.de
- Part 2: Kinds of RL Algorithms - Spinning Up documentationspinningup.openai.com
- [2005.01643] Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problemsar5iv.labs.arxiv.org
- pistar06.pdfpi.website
- Deep Reinforcement Learning Doesn't Work Yetalexirpan.com
- If it makes you feel any better, I've been doing this for a while and it took me... | Hacker Newsnews.ycombinator.com