Verbalizable Representations Form a Global Workspace in Language Models
If the mind is an ocean, we spend our lives floating at the surface. Beneath us, an enormous amount of processing takes place without our knowledge: our visual systems parsing the contours of a face, our motor circuits maintaining our posture. At any given moment, only a small fraction of this neural activity is accessible to us. Yet it is this privileged sliver of activity that we rely on to reason deliberately: to plan what ingredients to buy for a recipe, or to puzzle out why an engine won’t start. Such thoughts can be articulated out loud, deliberately held in mind, and brought to bear on whatever task the moment demands. This distinction, between our accessible thoughts and our unconscious processing, is perhaps the most striking feature of human cognition. In this paper, we present evidence that an analogous functional distinction has emerged in modern AI models. Specifically, we observe that language models maintain a privileged set of internal representations, available for rep
Verbalizable Representations Form a Global Workspace in Language Models Transformer Circuits Thread Verbalizable Representations Form a Global Workspace in Language Models Verbalizable Representations Form a Global Workspace in Language Models Authors Wes Gurnee * , Nicholas Sofroniew * Adam Pearce, Mateusz Piotrowski, Isaac Kauvar, Runjin Chen, Anna Soligo, Paul Bogdan, Euan Ong, Rowan Wang, Ben Thompson, David Abrahams, Subhash Kantamneni, Emmanuel Ameisen, Joshua Batson Jack Lindsey *† Affiliations Anthropic Published July 6, 2026 * Core contributor; † Correspondence to jacklindsey@anthropi
Explore this link on the map →saved by
- Linda
- Julian Quevedo
- Scott Langille
- Asher P
- Neel Redkar
- Lydia Nottingham
- Timothy Kostolansky
- Kushal Thaman
- [David L]
- Idhant Gulati
- 5514
- Yvonne Chen
related reading
- A global workspace in language models \ Anthropicanthropic.com
- No Space Like J-Space - by Zvi Mowshowitzthezvi.substack.com
- Language Models, World Models, and Human Model-Buildinglingo.csail.mit.edu
- A Review of Anthropic's Global Workspace Paper — LessWronglesswrong.com
- On the Biology of a Large Language Modeltransformer-circuits.pub
- Circuit Tracing: Revealing Computational Graphs in Language Modelstransformer-circuits.pub
- Tracing the thoughts of a large language model \ Anthropicanthropic.com
- Transformer Circuits Threadtransformer-circuits.pub
- Can activation verbalizers surface an internal chain of thought? — LessWronglesswrong.com
- Natural Language Autoencoders Produce Unsupervised Explanations of LLM Activationstransformer-circuits.pub
- Prediction, Explanation, or Over-interpretation?elena-baixy.github.io
- Jacobian Lens – Qwen3.6-27B | Neuronpedianeuronpedia.org