AGI safety from first principles: Introduction — LessWrong
This is the first part of a six-part report called AGI safety from first principles, in which I've attempted to put together the most complete and compelling case I can for why the development of AGI might pose an existential threat. The report stems from my dissatisfaction with existing arguments about the potential risks from AGI. Early work tends to be less relevant in the context of modern machine learning; more recent work is scattered and brief. I originally intended to just summarise other people's arguments, but as this report has grown, it's become more representative of my own views and less representative of anyone else's. So while it covers the standard ideas, I also think that it provides a new perspective on how to think about AGI - one which doesn't take any previous claims for granted, but attempts to work them out from first principles. Having said that, the breadth of the topic I'm attempting to cover means that I've included many arguments which are only hastily sket
x AGI safety from first principles: Introduction — LessWrong Alignment & Agency AI Safety Public Materials AI Curated 129 AGI safety from first principles: Introduction by Richard_Ngo 28th Sep 2020 AI Alignment Forum 3 min read 18 129 Ω 39 This is the first part of a six-part report called AGI safety from first principles, in which I've attempted to put together the most complete and compelling case I can for why the development of AGI might pose an existential threat. The report stems from my dissatisfaction with existing arguments about the potential risks from AGI. Early work tends to be le
Explore this link on the map →saved by
related reading
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- Existential risk from artificial intelligence - Wikipediaen.wikipedia.org
- Commentary on AGI Safety from First Principles — AI Alignment Forumalignmentforum.org
- Where I agree and disagree with Eliezer — LessWronglesswrong.com
- A Simple Explanation of AGI Risk — AI Alignment Forumalignmentforum.org
- Existential risk from artificial intelligence - Wikipediaen.wikipedia.org
- Planning for AGI and beyond | OpenAIopenai.com
- A Field Guide to AI Safety—Asteriskasteriskmag.com
- FAQ on Catastrophic AI Risks | Yoshua Bengioyoshuabengio.org
- Unfalsifiable stories of doom | Mechanize, Inc.mechanize.work
- Why AI risks are the world’s most pressing problems | 80,000 Hours80000hours.org