The Plan - 2025 Update — LessWrong
For several years now, around the end of the year, I (John) write a post on our plan for AI alignment. That plan hasn’t changed too much over the past few years, so both this year’s post and last year’s are written as updates to The Plan - 2023 Version. I’ll give a very quick outline here of what’s in the 2023 Plan post. If you have questions or want to argue about points, you should probably go to that post to get the full version. 2023 and 2024 were mostly focused on Natural Latents - we’ll talk more shortly about that work and how it fits into the bigger picture. In 2025, we did continue to put out some work on natural latents, but our main focus has shifted. Natural latents are a major foothold on understanding natural abstraction. One could reasonably argue that they’re the only rigorous foothold on the core problem to date, the first core mathematical piece of the future theory. We’ve used that foothold to pull ourselves up a bit, and can probably pull ourselves up a little furth
x The Plan - 2025 Update — LessWrong AI Frontpage 96 The Plan - 2025 Update by johnswentworth , David Lorell 31st Dec 2025 9 min read 21 96 What’s “The Plan”? For several years now, around the end of the year, I (John) write a post on our plan for AI alignment. That plan hasn’t changed too much over the past few years, so both this year’s post and last year’s are written as updates to The Plan - 2023 Version . I’ll give a very quick outline here of what’s in the 2023 Plan post. If you have questions or want to argue about points, you should probably go to that post to get the full version. Wha
Explore this link on the map →related reading
- The Plan - 2023 Version — LessWronglesswrong.com
- LessWronglesswrong.com
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- Shah and Yudkowsky on alignment failures — LessWronglesswrong.com
- On how various plans miss the hard bits of the alignment challenge — LessWronglesswrong.com
- A positive case for how we might succeed at prosaic AI alignment — AI Alignment Forumalignmentforum.org
- The Best of LessWrong — LessWronglesswrong.com
- The Field of AI Alignment: A Postmortem, and What To Do About It — LessWronglesswrong.com
- Alignment By Default — AI Alignment Forumalignmentforum.org
- The Iliad Intensive Course Materials — LessWronglesswrong.com
- What Is The Alignment Problem? — LessWronglesswrong.com