Limits to Legibility — LessWrong
From time to time, someone makes the case for why transparency in reasoning is important. The latest conceptualization is Epistemic Legibility by Eli…
x Limits to Legibility — LessWrong Epistemology Tacit Knowledge Rationality World Modeling Frontpage 168 Limits to Legibility by Jan_Kulveit 29th Jun 2022 6 min read 12 168 From time to time, someone makes the case for why transparency in reasoning is important. The latest conceptualization is Epistemic Legibility by Elizabeth, but the core concept is similar to reasoning transparency used by OpenPhil, and also has some similarity to A Sketch of Good Communication by Ben Pace. I'd like to offer a gentle pushback. The tl;dr is in my comment on Ben's post, but it seems useful enough for a standa
Explore this link on the map →related reading
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- Dario Amodei — The Urgency of Interpretabilitydarioamodei.com
- Legible vs. Illegible AI Safety Problems — LessWronglesswrong.com
- Where I agree and disagree with Eliezer — LessWronglesswrong.com
- How Go Players Disempower Themselves to AI — LessWronglesswrong.com
- the-illusion-of-thinking.pdfml-site.cdn-apple.com
- The Legibility Problempress.asimov.com
- Legible vs. Illegible AI Safety Problems — EA Forumforum.effectivealtruism.org
- Announcing ReasoningLens — Visualizing and Diagnosing LLM Reasoning at a Glancehuggingface.co
- [2606.20560] How Transparent is DiffusionGemma?arxiv.org
- Neel Nanda on Mechanistic Interpretability: Progress, Limits, and Paths to Safer AI — EA Forumforum.effectivealtruism.org
- RohanS's Shortform — LessWronglesswrong.com