Reading List - by Julian Stastny - Redwood Research blog
blog.redwoodresearch.org · 1,633 words · saved by 5 readers
(Last updated Jul 28th 2025)
(Last updated Jul 28th 2025) Section 1 is a quick guide to the key ideas in AI control, aimed at readers who want to get up to speed as quickly as possible. Section 2 is an extensive guide to almost all of our writing on AI risk, aimed at those who want to gain a deep understanding of Redwood’s worldview. The case for ensuring that powerful AIs are controlled (Buck Shlegeris and Ryan Greenblatt, May 2024): This is our original essay explaining why we think AI control is a valuable research direction. AI Catastrophes And Rogue Deployments (Buck Shlegeris, Jun 2024): We often refer to…
saved by
related reading
- AI Control: Improving Safety Despite Intentional Subversion — AI Alignment Forumalignmentforum.org
- The case for ensuring that powerful AIs are controlled — LessWronglesswrong.com
- The Case Against AI Control Research — LessWronglesswrong.com
- The case for ensuring that powerful AIs are controlledblog.redwoodresearch.org
- Reading Listredwoodresearch.substack.com
- AI Control: Improving Safety Despite Intentional Subversion — LessWronglesswrong.com
- Research Areas in AI Control (The Alignment Project by UK AISI) — AI Alignment Forumalignmentforum.org
- I think alignment work is more promising than control work — LessWronglesswrong.com
- Introduction to AI Control - by Sarah - BlueDot Impactblog.bluedot.org
- Thoughts on the conservative assumptions in AI controlblog.redwoodresearch.org
- 7+ tractable directions in AI control — AI Alignment Forumalignmentforum.org
- How useful is AI control? @ Trackstracks.xlabtracks.workers.dev