The Platonic Representation Hypothesis
This is experimental HTML to improve accessibility. We invite you to report rendering errors. Use Alt+Y to toggle on accessible reporting links and Alt+Shift+Y to toggle off. Learn more about this project and help improve conversions. HTML conversions sometimes display errors due to content that did not convert correctly from the source. This paper uses the following packages that are not yet supported by the HTML conversion tool. Feedback on these issues are not necessary; they are known and are being worked on. Authors: achieve the best HTML results from your LaTeX submissions by following these best practices. We argue that representations in AI models, particularly deep networks, are converging. First, we survey many examples of convergence in the literature: over time and across multiple domains, the ways by which different neural networks represent data are becoming more aligned. Next, we demonstrate convergence across data modalities: as vision models and language models get la
The Platonic Representation Hypothesis Minyoung Huh Brian Cheung Tongzhou Wang Phillip Isola Abstract We argue that representations in AI models, particularly deep networks, are converging. First, we survey many examples of convergence in the literature: over time and across multiple domains, the ways by which different neural networks represent data are becoming more aligned. Next, we demonstrate convergence across data modalities: as vision models and language models get larger, they measure distance between datapoints in a more and more alike way. We hypothesize that this convergence is dri
Explore this link on the map →related reading
- The Platonic Representation Hypothesisphillipi.github.io
- [2405.07987] The Platonic Representation Hypothesisarxiv.org
- All AI Models Might Be The Same - by Jack Morrisblog.jxmo.io
- The Scaling Hypothesis · Gwern.netgwern.net
- Verbalizable Representations Form a Global Workspace in Language Modelstransformer-circuits.pub
- Interpretability Dreamstransformer-circuits.pub
- On neural scaling and the quanta hypothesisericjmichaud.com
- Formation of Representations in Neural Networksarxiv.org
- [2210.13382] Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Taskarxiv.org
- Large Language Model: world models or surface statistics?thegradient.pub
- Learned feature representations are biased by complexity, learning order, position, and morearxiv.org
- 07-representation-learning.pdfweb.stanford.edu