Fields that I reference when thinking about AI takeover prevention — LessWrong
lesswrong.com · 4,885 words · saved by 2 readers
Is AI takeover like a nuclear meltdown? A coup? A plane crash? …
x Fields that I reference when thinking about AI takeover prevention — LessWrong AI Curated 146 Fields that I reference when thinking about AI takeover prevention by Buck 13th Aug 2024 AI Alignment Forum Linkpost for redwoodresearch.substack.com 12 min read 16 146 Ω 59 Is AI takeover like a nuclear meltdown? A coup? A plane crash? My day job is thinking about safety measures that aim to reduce catastrophic risks from AI (especially risks from egregious misalignment). The two main themes of this work are the design of such measures (what’s the space of techniques we might expect to be affordabl
saved by
related reading
- How might we safely pass the buck to AI? — LessWronglesswrong.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- What failure looks like — LessWronglesswrong.com
- The case for ensuring that powerful AIs are controlled — LessWronglesswrong.com
- The Case Against AI Control Research — LessWronglesswrong.com
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- Reading Listblog.redwoodresearch.org
- AI Control: Improving Safety Despite Intentional Subversion — LessWronglesswrong.com
- AI Control: Improving Safety Despite Intentional Subversion — AI Alignment Forumalignmentforum.org
- gdm-ai-control-roadmap.pdfstorage.googleapis.com
- The case for ensuring that powerful AIs are controlledblog.redwoodresearch.org