Plans A, B, C, and D for misalignment risk — LessWrong
I sometimes think about plans for how to handle misalignment risk. Different levels of political will for handling misalignment risk result in differ…
x Plans A, B, C, and D for misalignment risk — LessWrong AI Frontpage 2025 Top Fifty: 19 % 139 Plans A, B, C, and D for misalignment risk by ryan_greenblatt 8th Oct 2025 AI Alignment Forum 8 min read 77 139 Ω 61 I sometimes think about plans for how to handle misalignment risk. Different levels of political will for handling misalignment risk result in different plans being the best option. I often divide this into Plans A, B, C, and D (from most to least political will required). See also Buck's quick take about different risk level regimes . In this post, I'll explain the Plan A/B/C/D abstra
Explore this link on the map →saved by
related reading
- Short AI Timelines Aren’t Always Higher-Leverageforethought.org
- What failure looks like — LessWronglesswrong.com
- The Case Against AI Control Research — LessWronglesswrong.com
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- Introducing Plan A - by Scott Alexander - Astral Codex Tenastralcodexten.com
- AI 2040: Plan A — LessWronglesswrong.com
- Ten people on the inside - by Buck Shlegerisredwoodresearch.substack.com
- What’s the short timeline plan? — LessWronglesswrong.com
- The Three Filters: Why Almost Every Plan to Survive ASI Fails Miserably — LessWronglesswrong.com
- You will be OK — LessWronglesswrong.com
- I Would Have Solved Alignment, But I Was Worried That Would Advance Timelines — LessWronglesswrong.com
- Planning for Extreme AI Risks — AI Alignment Forumalignmentforum.org