Statistical Physics for Ambitious Interpretability: A Workshop Retrospective — LessWrong
In April, we held our second workshop on statistical physics and AI safety, organized around a research agenda we've been building over the past year…
x Statistical Physics for Ambitious Interpretability: A Workshop Retrospective — LessWrong AI Frontpage 5 Statistical Physics for Ambitious Interpretability: A Workshop Retrospective by Lauren Greenspan , Lucas Teixeira , ClaudineLim 12th Jun 2026 7 min read 0 5 In April, we held our second workshop on statistical physics and AI safety, organized around a research agenda we've been building over the past year: Physics Inspired Ambitious Mechanistic Interpretability (PIAMI). ( Stay tuned for a PIAMI research roadmap, to be shared soon. ) This post is a summary of the week's talks, as well as my
saved by
related reading
- Neel Nanda on Mechanistic Interpretability: Progress, Limits, and Paths to Safer AI — EA Forumforum.effectivealtruism.org
- A Pragmatic Vision for Interpretability — LessWronglesswrong.com
- Dario Amodei — The Urgency of Interpretabilitydarioamodei.com
- Problem Areas in Physics and AI Safety | Apart Researchapartresearch.com
- An Ambitious Vision for Interpretability — AI Alignment Forumalignmentforum.org
- A Pragmatic Vision for Interpretability — AI Alignment Forumalignmentforum.org
- Intentionally Designing the Future of AIgoodfire.ai
- Assessing skeptical views of interpretability research | Christopher Pottsweb.stanford.edu
- On Optimism for Interpretabilitygoodfire.ai
- What is the purpose of interpretability?ericjmichaud.com
- How To Become A Mechanistic Interpretability Researcher — AI Alignment Forumalignmentforum.org
- The Hitchhiker's Guide to Actionable Interpretabilityactionable-interpretability-guide.github.io