Language Models, World Models, and Human Model-Building
Does the computation performed by language models involve real language understanding, or just string manipulation? When the first really high-quality neural language models first appeared, this Big Question was mainly formulated in terms of whether LMs could represent meaning. The general recipe was to inspect an LM’s internal representations, and to find a simple procedure for translating them into something resembling a human-approved meaning representation in some formal linguistic theory. It turned out that the hard part of this problem had nothing to do with LMs: it was figuring out what “meaning” means, and what these human-approved representations should look like in the first place. So the NLP interpretability community immediately started arguing about model-theoretic vs relational theories of meaning; whether it was enough for models’ internal representations to be isomorphic to human-designed meaning representations, or whether explicit links to perception and action were s
Language Models, World Models, and Human Model-Building Language Models, World Models, and Human Model-Building Jacob Andreas, 26 July 2024 Does the computation performed by language models involve real language understanding, or just string manipulation? When the first really high-quality neural language models first appeared, this Big Question was mainly formulated in terms of whether LMs could represent meaning . The general recipe was to inspect an LM’s internal representations, and to find a simple procedure for translating them into something resembling a human-approved meaning represent
Explore this link on the map →saved by
related reading
- Verbalizable Representations Form a Global Workspace in Language Modelstransformer-circuits.pub
- LLMs and World Models, Part 1 - by Melanie Mitchellaiguide.substack.com
- Transformer Circuits Threadtransformer-circuits.pub
- Large Language Model: world models or surface statistics?thegradient.pub
- World Models: Computing the Uncomputablenotboring.co
- Tracing the thoughts of a large language model \ Anthropicanthropic.com
- Prediction, Explanation, or Over-interpretation?elena-baixy.github.io
- Do large language models understand us? | by Blaise Aguera y Arcas | Mediummedium.com
- Physics of Language Modelsphysics.allen-zhu.com
- [2210.13382] Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Taskarxiv.org
- Language models can explain neurons in language modelsopenaipublic.blob.core.windows.net
- Language Modelinglena-voita.github.io