The prototypical catastrophic AI action is getting root access to its datacenter - AI Alignment Forum
(I think Carl Shulman came up with the “hacking the SSH server” example, thanks to him for that. Thanks to Ryan Greenblatt, Jenny Nitishinskaya, and Ajeya Cotra for comments.) …
x The prototypical catastrophic AI action is getting root access to its datacenter — AI Alignment Forum Existential risk AI Frontpage 68 The prototypical catastrophic AI action is getting root access to its datacenter by Buck 2nd Jun 2022 3 min read 13 68 (I think Carl Shulman came up with the “hacking the SSH server” example, thanks to him for that. Thanks to Ryan Greenblatt, Jenny Nitishinskaya, and Ajeya Cotra for comments.) EDIT: I recommend reading my discussion with Oli in the comments for various useful clarifications. In my opinion, the prototypical example of an action which an AI can
saved by
related reading
- The prototypical catastrophic AI action is getting root access to its datacenter — LessWronglesswrong.com
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- The case for ensuring that powerful AIs are controlled — LessWronglesswrong.com
- What failure looks like — LessWronglesswrong.com
- What failure looks like — AI Alignment Forumalignmentforum.org
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- AI catastrophes and rogue deployments - by Buck Shlegerisblog.redwoodresearch.org
- AI Control: Improving Safety Despite Intentional Subversion — AI Alignment Forumalignmentforum.org
- Fields that I reference when thinking about AI takeover prevention — LessWronglesswrong.com
- FAQ on Catastrophic AI Risks | Yoshua Bengioyoshuabengio.org
- AI Control: Improving Safety Despite Intentional Subversion — LessWronglesswrong.com
- Buck Shlegeris on controlling AI that wants to take over – so we can use it anyway | 80,000 Hours80000hours.org