Commentary on AGI Safety from First Principles — AI Alignment Forum
My AGI safety from first principles report (which is now online here) was originally circulated as a google doc. Since there was a lot of good discussion in comments on the original document, I thought it would be worthwhile putting some of it online, and have copied out most of the substantive comment threads here. Many thanks to all of the contributors for their insightful points, and to Habryka for helping with formatting. Note that in some cases comments may refer to parts of the report that didn't make it into the public version. Will MacAskill Thanks so much for writing this! Huge +1 to more foundational work in this area. My overall biggest worry with your argument is just whether it's spending a lot of time defending something that's not really where the controversy lies. (This is true for me; I don't know if I'm idiosyncratic.) Distinguish two claims one could argue for: Claim 1: At some point in the future, assuming continued tech progress, history will have primarily become
x Commentary on AGI Safety from First Principles — AI Alignment Forum AGI safety from first principles AI Curated 34 Commentary on AGI Safety from First Principles by Richard_Ngo 23rd Nov 2020 65 min read 4 34 My AGI safety from first principles report (which is now online here ) was originally circulated as a google doc. Since there was a lot of good discussion in comments on the original document, I thought it would be worthwhile putting some of it online, and have copied out most of the substantive comment threads here. Many thanks to all of the contributors for their insightful points, and
Explore this link on the map →related reading
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- AGI safety from first principles: Introduction — LessWronglesswrong.com
- Planning for AGI and beyond | OpenAIopenai.com
- A positive case for how we might succeed at prosaic AI alignment — AI Alignment Forumalignmentforum.org
- Where I agree and disagree with Eliezer — LessWronglesswrong.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Optimality is the tiger, and agents are its teeth — LessWronglesswrong.com
- AGI safety career advice — EA Forumforum.effectivealtruism.org
- Unfalsifiable stories of doom | Mechanize, Inc.mechanize.work
- My take on Vanessa Kosoy's take on AGI safety — AI Alignment Forumalignmentforum.org
- My take on Jacob Cannell’s take on AGI safety — LessWronglesswrong.com