✳flâneur — a map of the web's best reading
A Theory of Deep Learning | Elements of a Vector Space
elonlit.com · 1,837 words · saved by 6 readers
We finally know why deep learning works.
Borges wrote a story about a man named Funes who, after a horseback accident, acquires the ability to perceive and remember everything. Every leaf on every tree. Every ripple on every stream at every moment. He is the perfect empiricist. Infinite data, infinite recall, infinite resolution. And he cannot think. Because thinking, as Borges understood, requires forgetting. Funes could reconstruct entire days from memory but could not understand why the dog at 3:14, seen from the side, should be called the same thing as the dog at 3:15, seen from the front. I suspect [that Funes] was not very good
Explore this link on the map →saved by
related reading
- Elon Litman | Elements of a Vector Spaceelonlit.com
- [2605.01172] A Theory of Generalization in Deep Learningarxiv.org
- [2604.21691] There Will Be a Scientific Theory of Deep Learningarxiv.org
- Maybe I was too harsh on deep learning theory (three days ago) — LessWronglesswrong.com
- The Little Book of Deep Learningfleuret.org
- The Decade of Deep Learning | Leo Gaobmk.sh
- Zipfian grokking | Jasper Gilleyjagilley.github.io
- Clare Lyle | What's grokking good for?clarelyle.com
- NL.pdfabehrouz.github.io
- Some Math behind Neural Tangent Kernel | Lil'Loglilianweng.github.io
- Statistical Mechanics of Deep Learningganguli-gang.stanford.edu
- [2503.02113] Deep Learning is Not So Mysterious or Differentarxiv.org