On the Biology of a Large Language Model
We investigate the internal mechanisms used by Claude 3.5 Haiku — Anthropic's lightweight production model — in a variety of contexts, using our circuit tracing methodology.
On the Biology of a Large Language Model × Transformer Circuits Thread On the Biology of a Large Language Model On the Biology of a Large Language Model We investigate the internal mechanisms used by Claude 3.5 Haiku — Anthropic's lightweight production model — in a variety of contexts, using our circuit tracing methodology. Authors Jack Lindsey † , Wes Gurnee * , Emmanuel Ameisen * , Brian Chen * , Adam Pearce * , Nicholas L. Turner * , Craig Citro * , David Abrahams, Shan Carter, Basil Hosmer, Jonathan Marcus, Michael Sklar, Adly Templeton, Trenton Bricken, Callum McDougall ◊ , Hoagy Cunning
Explore this link on the map →saved by
- Winnie Xu
- Trang Đoàn
- abinaya dinesh
- Tasha Pais
- Ratan Kaliani
- Freeman Jiang
- Aryan Naik
- Samuel Lo
- Matthew Wang
- Asher P
- Vyom Pathak
- Lydia Nottingham
related reading
- Tracing the thoughts of a large language model \ Anthropicanthropic.com
- Circuit Tracing: Revealing Computational Graphs in Language Modelstransformer-circuits.pub
- Transformer Circuits Threadtransformer-circuits.pub
- Mapping the Mind of a Large Language Model \ Anthropicanthropic.com
- Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnettransformer-circuits.pub
- A global workspace in language models \ Anthropicanthropic.com
- Tracing the Thoughts of a Large Language Model — LessWronglesswrong.com
- Tracing the thoughts of a large language model \ Anthropicanthropic.com
- Verbalizable Representations Form a Global Workspace in Language Modelstransformer-circuits.pub
- Natural Language Autoencoders \ Anthropicanthropic.com
- An Extremely Opinionated Annotated List of My Favourite Mechanistic Interpretability Papers v2 — AI Alignment Forumalignmentforum.org
- Emotion concepts and their function in a large language model \ Anthropicanthropic.com