MATS Applications + Research Directions I'm Currently Excited About — AI Alignment Forum
I've just opened summer MATS applications (where I'll supervise people to write mech interp papers) I'd love to get applications from any readers who…
x MATS Applications + Research Directions I'm Currently Excited About — AI Alignment Forum Interpretability (ML & AI) MATS Program AI Frontpage 31 MATS Applications + Research Directions I'm Currently Excited About by Neel Nanda 6th Feb 2025 9 min read 7 31 I've just opened summer MATS applications (where I'll supervise people to write mech interp papers) I'd love to get applications from any readers who are interested! Apply here , due Feb 28 As part of this, I wrote up a list of research areas I'm currently excited about, and thoughts for promising directions within those, which I thought mi
Explore this link on the map →related reading
- An Extremely Opinionated Annotated List of My Favourite Mechanistic Interpretability Papers v2 — AI Alignment Forumalignmentforum.org
- Transformer Circuits Threadtransformer-circuits.pub
- Dario Amodei — The Urgency of Interpretabilitydarioamodei.com
- An Extremely Opinionated Annotated List of My Favourite Mechanistic Interpretability Papers v2 — AI Alignment Forumalignmentforum.org
- Tracing the thoughts of a large language model \ Anthropicanthropic.com
- How To Become A Mechanistic Interpretability Researcher — AI Alignment Forumalignmentforum.org
- A Pragmatic Vision for Interpretability — AI Alignment Forumalignmentforum.org
- How Can Interpretability Researchers Help AGI Go Well? — AI Alignment Forumalignmentforum.org
- Neel Nanda on Mechanistic Interpretability: Progress, Limits, and Paths to Safer AI — EA Forumforum.effectivealtruism.org
- An Intuitive Explanation of Sparse Autoencoders for LLM Interpretability | Adam Karvonenadamkarvonen.github.io
- An Ambitious Vision for Interpretability — AI Alignment Forumalignmentforum.org
- How To Become A Mechanistic Interpretability Researcher — LessWronglesswrong.com