The Case Against AI Control Research — LessWrong
lesswrong.com · 11,727 words · saved by 6 readers
The AI Control Agenda, in its own words: …
x The Case Against AI Control Research — LessWrong AI Control 2023 Longform Reviews AI Curated 2025 Top Fifty: 65 % 433 The Case Against AI Control Research by johnswentworth 21st Jan 2025 8 min read 85 433 The AI Control Agenda, in its own words : … we argue that AI labs should ensure that powerful AIs are controlled . That is, labs should make sure that the safety measures they apply to their powerful models prevent unacceptably bad outcomes, even if the AIs are misaligned and intentionally try to subvert those safety measures. We think no fundamental research breakthroughs are required for
saved by
related reading
- The case for ensuring that powerful AIs are controlled — LessWronglesswrong.com
- Reading Listblog.redwoodresearch.org
- I think alignment work is more promising than control work — LessWronglesswrong.com
- Thoughts on the conservative assumptions in AI controlblog.redwoodresearch.org
- Introduction to AI Control - by Sarah - BlueDot Impactblog.bluedot.org
- AI 2040: Plan Aai-2040.com
- AI Control: Improving Safety Despite Intentional Subversion — AI Alignment Forumalignmentforum.org
- The case for ensuring that powerful AIs are controlledblog.redwoodresearch.org
- What failure looks like — AI Alignment Forumalignmentforum.org
- How can we solve diffuse threats like research sabotage with AI control?blog.redwoodresearch.org
- How useful is AI control? @ Trackstracks.xlabtracks.workers.dev
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com