[2508.14802] Privileged Self-Access Matters for Introspection in AI
arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website. Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them. Have an idea for a project that will add value for arXiv's community? Learn more about arXivLabs. arXiv Operational Status
[2508.14802] Privileged Self-Access Matters for Introspection in AI Skip to main content arXiv is now an independent nonprofit! Learn more × Search arXiv Press Enter to search · Advanced search --> Computer Science > Artificial Intelligence arXiv:2508.14802 (cs) [Submitted on 20 Aug 2025] Title: Privileged Self-Access Matters for Introspection in AI Authors: Siyuan Song , Harvey Lederman , Jennifer Hu , Kyle Mahowald View a PDF of the paper titled Privileged Self-Access Matters for Introspection in AI, by Siyuan Song and 3 other authors View PDF HTML (experimental) Abstract: Wheth
Explore this link on the map →related reading
- [2508.14802] Privileged Self-Access Matters for Introspection in AIarxiv.org
- [2410.13787] Looking Inward: Language Models Can Learn About Themselves by Introspectionarxiv.org
- LLMs can learn about themselves by introspection — LessWronglesswrong.com
- [2410.13787] Looking Inward: Language Models Can Learn About Themselves by Introspectionarxiv.org
- [2604.16812] Introspection Adapters: Training LLMs to Report Their Learned Behaviorsarxiv.org
- Emergent introspective awareness in large language models \ Anthropicanthropic.com
- [2506.05068] Does It Make Sense to Speak of Introspection in Large Language Models?arxiv.org
- Introspection Adapters: Training LLMs to Report Their Learned Behaviorsarxiv.org
- A Pragmatic Vision for Interpretability — AI Alignment Forumalignmentforum.org
- Small Models Can Introspect, Toovgel.me
- Emergent Introspective Awareness in Large Language Modelstransformer-circuits.pub
- [2601.01828] Emergent Introspective Awareness in Large Language Modelsarxiv.org