If anyone builds it, everyone will plausibly be fine - LessWrong 2.0 viewer
I think AI takeover is plausible. But Eliezer’s argument that it’s more than 98% likely to happen does not stand up to scrutiny, and I’m worried that MIRI’s overconfidence has reduced the credibility of the issue. Here is why I think the core argument in "if anyone builds it, everyone dies" is much weaker than the authors claim.This post was written in a personal capacity. Most of this content has been written-up before by a combination of Paul Christiano, Joe Carlsmith, and others. But to my knowledge, this content has not yet been consolidated into a direct response to MIRI’s core case for alignment difficulty.
If anyone builds it, everyone will plausibly be fine - LessWrong 2.0 viewer If anyone builds it, everyone will plausibly be fine joshc 18 Sep 2025 20:03 UTC LW: 32 AF: 18 24 comments 7 min read LW link IABIED AI I think AI takeover is plausible. But Eliezer’s argument that it’s more than 98% likely to happen does not stand up to scrutiny, and I’m worried that MIRI’s overconfidence has reduced the credibility of the issue. Here is why I think the core argument in “if anyone builds it, everyone dies” is much weaker than the authors claim. This post was written in a personal capacity. Most of thi
Explore this link on the map →related reading
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- What failure looks like — LessWronglesswrong.com
- Where I agree and disagree with Eliezer — LessWronglesswrong.com
- Current AIs seem pretty misaligned to me — LessWronglesswrong.com
- Unfalsifiable stories of doom | Mechanize, Inc.mechanize.work
- A positive case for how we might succeed at prosaic AI alignment — AI Alignment Forumalignmentforum.org
- Current AIs seem pretty misaligned to meblog.redwoodresearch.org
- Another (outer) alignment failure story — AI Alignment Forumalignmentforum.org
- Plans A, B, C, and D for misalignment risk — LessWronglesswrong.com
- Shah and Yudkowsky on alignment failures — LessWronglesswrong.com
- The Problem — LessWronglesswrong.com