✳flâneur — a map of the web's best reading
arxiv.org/pdf/1805.08522
arxiv.org · 14,767 words · saved by 2 readers
N/A
# link_12qr17hk9mx.pdf ## Metadata - PDFFormatVersion=1.5 - IsLinearized=false - IsAcroFormPresent=false - IsXFAPresent=false - IsCollectionPresent=false - IsSignaturesPresent=false - CreationDate=D:20190423011718Z - Creator=LaTeX with hyperref package - ModDate=D:20190423011718Z - Custom.PTEX.Fullbanner=This is pdfTeX, Version 3.14159265-2.6-1.40.17 (TeX Live 2016) kpathsea version 6.2.2 - Producer=pdfTeX-1.40.17 - Trapped=False ## Contents ### Page 1 Published as a conference paper at ICLR 2019DEEP LEARNING GENERALIZES BECAUSE THE PARAMETER-FUNCTION MAP IS BIASED TOWARDS SIMPLE FUNCTIONSG
Explore this link on the map →saved by
related reading
- The generalization phase diagram — LessWronglesswrong.com
- [1805.08522] Deep learning generalizes because the parameter-function map is biased towards simple functionsarxiv.org
- Bayesian Neural Networkscs.toronto.edu
- [2503.02113] Deep Learning is Not So Mysterious or Differentarxiv.org
- Maybe I was too harsh on deep learning theory (three days ago) — LessWronglesswrong.com
- Some Math behind Neural Tangent Kernel | Lil'Loglilianweng.github.io
- The Scaling Hypothesis · Gwern.netgwern.net
- Neural networks generalize because of this one weird trick — LessWronglesswrong.com
- A Theory of Deep Learning | Elements of a Vector Spaceelonlit.com
- [1802.05296] Stronger generalization bounds for deep nets via a compression approacharxiv.org
- Neural networks generalize because of this one weird trick — AI Alignment Forumalignmentforum.org
- Statistical Mechanics of Deep Learningganguli-gang.stanford.edu