The prototypical catastrophic AI action is getting root access to its datacenter - AI Alignment Forum
(I think Carl Shulman came up with the “hacking the SSH server” example, thanks to him for that. Thanks to Ryan Greenblatt, Jenny Nitishinskaya, and Ajeya Cotra for comments.) …
x The prototypical catastrophic AI action is getting root access to its datacenter — AI Alignment Forum Existential risk AI Frontpage 68 The prototypical catastrophic AI action is getting root access to its datacenter by Buck 2nd Jun 2022 3 min read 13 68 (I think Carl Shulman came up with the “hacking the SSH server” example, thanks to him for that. Thanks to Ryan Greenblatt, Jenny Nitishinskaya, and Ajeya Cotra for comments.) EDIT: I recommend reading my discussion with Oli in the comments for various useful clarifications. In my opinion, the prototypical example of an action which an AI can
Explore this link on the map →saved by
related reading
- The prototypical catastrophic AI action is getting root access to its datacenter — LessWronglesswrong.com
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- The case for ensuring that powerful AIs are controlled — LessWronglesswrong.com
- What failure looks like — LessWronglesswrong.com
- AI catastrophes and rogue deployments - by Buck Shlegerisblog.redwoodresearch.org
- Fields that I reference when thinking about AI takeover prevention — LessWronglesswrong.com
- FAQ on Catastrophic AI Risks | Yoshua Bengioyoshuabengio.org
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- Buck Shlegeris on controlling AI that wants to take over – so we can use it anyway | 80,000 Hours80000hours.org
- AI Control: Improving Safety Despite Intentional Subversion — AI Alignment Forumalignmentforum.org
- AI Control: Improving Safety Despite Intentional Subversion — LessWronglesswrong.com
- A basic systems architecture for AI agents that do autonomous research — LessWronglesswrong.com