✳flâneur — a map of the web's best reading
A case for LLMs as Self-predictors — LessWrong
lesswrong.com · 4,753 words · saved by 1 readers
Written as part of the MATS 9.1 extension program, mentored by Richard Ngo. Additional thanks to Maria Kostylew for helpful draft feedback. …
x A case for LLMs as Self-predictors — LessWrong Agent Foundations Game Theory Language Models (LLMs) MATS Program AI Frontpage 32 A case for LLMs as Self-predictors by Ashe Vazquez Nuñez 5th Jul 2026 12 min read 6 32 Written as part of the MATS 9.1 extension program, mentored by Richard Ng o. Additional thanks to Maria Kostylew for helpful draft feedback. Introduction This post advocates a perspective of LLMs as seeking to minimise prediction error with respect to their world models. We can moreover interpret token outputs and their scaffolded consequences as actions that close a control loop
Explore this link on the map →related reading
- The Persona Selection Model: Why AI Assistants might Behave like Humansalignment.anthropic.com
- Simulators — LessWronglesswrong.com
- Guardian Angels: LLM Personalization for Productivity and Security · Gwern.netgwern.net
- The Waluigi Effect (mega-post) — LessWronglesswrong.com
- LLM Powered Autonomous Agents | Lil'Loglilianweng.github.io
- Role-playing vs Self-modelling — LessWronglesswrong.com
- trees are harlequins, words are harlequins - the voidnostalgebraist.tumblr.com
- LLMs can learn about themselves by introspection — LessWronglesswrong.com
- the void — LessWronglesswrong.com
- On the functional self of LLMs — LessWronglesswrong.com
- llm assistant personas seem increasingly incoherent (some subjective observations) — LessWronglesswrong.com
- The Artificial Self — LessWronglesswrong.com