✳flâneur — a map of the web's best reading
The Case Against AI Control Research — LessWrong
lesswrong.com · 11,727 words · saved by 4 readers
The AI Control Agenda, in its own words: …
x The Case Against AI Control Research — LessWrong AI Control 2023 Longform Reviews AI Curated 2025 Top Fifty: 65 % 433 The Case Against AI Control Research by johnswentworth 21st Jan 2025 8 min read 85 433 The AI Control Agenda, in its own words : … we argue that AI labs should ensure that powerful AIs are controlled . That is, labs should make sure that the safety measures they apply to their powerful models prevent unacceptably bad outcomes, even if the AIs are misaligned and intentionally try to subvert those safety measures. We think no fundamental research breakthroughs are required for
Explore this link on the map →saved by
related reading
- The case for ensuring that powerful AIs are controlled — LessWronglesswrong.com
- AI Control: Improving Safety Despite Intentional Subversion — AI Alignment Forumalignmentforum.org
- Thoughts on the conservative assumptions in AI controlblog.redwoodresearch.org
- How can we solve diffuse threats like research sabotage with AI control?blog.redwoodresearch.org
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- The case for ensuring that powerful AIs are controlledblog.redwoodresearch.org
- What failure looks like — LessWronglesswrong.com
- How might we safely pass the buck to AI? — LessWronglesswrong.com
- Introduction to AI Control - by Sarah - BlueDot Impactbluedot.org
- Fields that I reference when thinking about AI takeover prevention — LessWronglesswrong.com
- Plans A, B, C, and D for misalignment risk — LessWronglesswrong.com
- Introduction to AI Control - by Sarah - BlueDot Impactblog.bluedot.org