You can, in fact, bamboozle an unaligned AI into sparing your life — LessWrong
There has been a renewal of discussion on how much hope we should have of an unaligned AGI leaving humanity alive on Earth after a takeover. When this topic is discussed, the idea of using simulation arguments or acausal trade to make the AI spare our lives often come up. These ideas have a long history. The first mention I know of comes from Rolf Nelson in 2007 on an SL4 message board, the idea later makes a brief appearance in Superintelligence under the name of Anthropic Capture, and came up on LessWrong last time as recently as a few days ago. In response to these, Nate Soares wrote Decision theory does not imply that we get to have nice things, arguing that decision theory is not going to save us, and that we can't bamboozle a superintelligence into submission by clever simulation arguments. However, none of the posts I found so far on the topic present the strongest version of the argument, and while Nate Soares validly argues against various weaker versions, he doesn't address
x You can, in fact, bamboozle an unaligned AI into sparing your life — LessWrong Simulation Hypothesis Decision theory Existential risk AI Frontpage 127 You can, in fact, bamboozle an unaligned AI into sparing your life by David Matolcsi 29th Sep 2024 32 min read 173 127 There has been a renewal of discussion on how much hope we should have of an unaligned AGI leaving humanity alive on Earth after a takeover. When this topic is discussed, the idea of using simulation arguments or acausal trade to make the AI spare our lives often come up. These ideas have a long history. The first mention I kn
Explore this link on the map →related reading
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- The Hour I First Believed | Slate Star Codexslatestarcodex.com
- Making deals with early schemers — LessWronglesswrong.com
- Are You Living in a Computer Simulation?simulation-argument.com
- Cyborgism — LessWronglesswrong.com
- The case against AI alignment — LessWronglesswrong.com
- A positive case for how we might succeed at prosaic AI alignment — AI Alignment Forumalignmentforum.org
- My AI Opinions - by Scott Alexander - Astral Codex Tensubstack.com
- Unfalsifiable stories of doom | Mechanize, Inc.mechanize.work
- AI Could Defeat All Of Us Combined — LessWronglesswrong.com
- ⿻ Symbiogenesis vs. Convergent Consequentialism — LessWronglesswrong.com
- Are You Living in a Simulation?simulation-argument.com