Where I agree and disagree with Eliezer — LessWrong
Paul writes a list of 19 important places where he agrees with Eliezer on AI existential risk and safety, and a list of 27 places where he disagrees.…
x Where I agree and disagree with Eliezer — LessWrong Best of LessWrong 2022 AI Risk AI World Modeling Curated 933 Where I agree and disagree with Eliezer by paulfchristiano 19th Jun 2022 AI Alignment Forum 22 min read 224 933 Ω 218 ( Partially in response to AGI Ruin: A list of Lethalities . Written in the same rambling style. Not exhaustive. ) Agreements Powerful AI systems have a good chance of deliberately and irreversibly disempowering humanity. This is a much more likely failure mode than humanity killing ourselves with destructive physical technologies. Catastrophically risky AI systems
Explore this link on the map →saved by
related reading
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- Critical review of Christiano's disagreements with Yudkowsky — LessWronglesswrong.com
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- You will be OK — LessWronglesswrong.com
- My AI Opinions - by Scott Alexander - Astral Codex Tensubstack.com
- Unfalsifiable stories of doom | Mechanize, Inc.mechanize.work
- A Field Guide to AI Safety—Asteriskasteriskmag.com
- The Best of LessWrong — LessWronglesswrong.com
- Focus on the places where you feel shocked everyone's dropping the ball — LessWronglesswrong.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- A positive case for how we might succeed at prosaic AI alignment — AI Alignment Forumalignmentforum.org
- Notes on Existential Risk from Artificial Superintelligencemichaelnotebook.com