Many arguments for AI x-risk are wrong — LessWrong
lesswrong.com · 15,325 words · saved by 1 readers
The following is a lightly edited version of a memo I wrote for a retreat. It was inspired by a draft of Counting arguments provide no evidence for A…
x Many arguments for AI x-risk are wrong — LessWrong Deceptive Alignment AI Risk Skepticism AI Governance Counting arguments Language Models (LLMs) AI Community Frontpage 182 Many arguments for AI x-risk are wrong by TurnTrout 5th Mar 2024 AI Alignment Forum 15 min read 96 182 Ω 48 The following is a lightly edited version of a memo I wrote for a retreat. It was inspired by a draft of Counting arguments provide no evidence for AI doom . I think that my post covers important points not made by the published version of that post. I'm also thankful for the dozens of interesting conversations and
related reading
- Many arguments for AI x-risk are wrong — AI Alignment Forumalignmentforum.org
- AI in 2025: gestalt — LessWronglesswrong.com
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- Unfalsifiable stories of doom | Mechanize, Inc.mechanize.work
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- How will we update about scheming?blog.redwoodresearch.org
- Where I agree and disagree with Eliezer — LessWronglesswrong.com
- Existential Risk from AI: An Exposition for Mathematiciansalkjash.github.io
- The Best of LessWrong — LessWronglesswrong.com
- New report: “Scheming AIs: Will AIs fake alignment during training in order to get power?”joecarlsmith.com
- A positive case for how we might succeed at prosaic AI alignment — AI Alignment Forumalignmentforum.org