Bellman's Curse on Advice - dust-nib
In dynamic programming and reinforcement learning, the most fundamental challenge we face is just grokking the high dimensionality of a state space. Go is a ...
In dynamic programming and reinforcement learning, the most fundamental challenge we face is just grokking the high dimensionality of a state space. Go is a “simple” game whose rules can be fully explained on a single sheet of paper, yet those rules yield more board configurations than the number of atoms in the universe (~10174). Given the state space, how does any solver for this game even represent these configurations, let alone evaluate them? While games are helpful toys to start with, this phenomenon of “exponential explosion” in systems is extremely common in many real-world problems,…
related reading
- Curse of dimensionality - Wikipediaen.wikipedia.org
- Advice That Actually Worked For Me — Nabeel S. Qureshinabeelqu.co
- Evolution as Backstop for Reinforcement Learning · Gwern.netgwern.net
- Shtetl-Optimized >> Blog Archive >> The First Law of Complexodynamicsscottaaronson.blog
- Study Guide — LessWronglesswrong.com
- Debugging Reinforcement Learning Systemsandyljones.com
- Adviceansonyu.me
- c21f4ce780c5c9d774f79841b81fdc6d-Paper.pdfproceedings.neurips.cc
- How To Think Real Good | Meta-rationalitymetarationality.com
- Memorized Rules: How to give your life directionjulian.com
- 2203.00543arxiv.org
- go-explore-nature.pdfadrien.ecoffet.com