✳flâneur — a map of the web's best reading
Role-playing vs Self-modelling — LessWrong
lesswrong.com · 1,827 words · saved by 2 readers
In a recent debate on Twitter – which I recommend reading in full – David Chalmers argues: …
x Role-playing vs Self-modelling — LessWrong Language Models (LLMs) LLM Personas AI Frontpage 63 Role-playing vs Self-modelling by Jan_Kulveit 7th Apr 2026 4 min read 3 63 In a recent debate on Twitter – which I recommend reading in full – David Chalmers argues : "Claude doesn't role-play the assistant, it realizes the assistant. Role-playing and realization are quite distinct phenomena, even at the level of behavior and function." Jack Lindsey questions this , pointing out evidence in the opposite direction: "I'm curious what you'd say it's doing when it's sampling tokens on the user turn, or
Explore this link on the map →saved by
related reading
- The Persona Selection Model: Why AI Assistants might Behave like Humansalignment.anthropic.com
- trees are harlequins, words are harlequins - the voidnostalgebraist.tumblr.com
- Simulators — LessWronglesswrong.com
- llm assistant personas seem increasingly incoherent (some subjective observations) — LessWronglesswrong.com
- Discovering Language Model Behaviors with Model-Written Evaluations — LessWronglesswrong.com
- Guardian Angels: LLM Personalization for Productivity and Security · Gwern.netgwern.net
- The persona selection model — LessWronglesswrong.com
- The Artificial Selftheartificialself.ai
- On the functional self of LLMs — LessWronglesswrong.com
- The Waluigi Effect (mega-post) — LessWronglesswrong.com
- A Three-Layer Model of LLM Psychology — LessWronglesswrong.com
- The Artificial Self — LessWronglesswrong.com