Wei Dai's Shortform — AI Alignment Forum
There's a LW post titled Moltbook and the AI Alignment Problem but it seems unrelated to the question I'm interested in here. I'm trying to coordinate with, or avoid interfering with, people who are trying to implement an AI pause or create conditions conducive to a future pause. As mentioned in the grandparent comment, one way people like us could interfere with such efforts is by feeding into a human tendency to be overconfident about one's own ideas/solutions/approaches. Personal experience with Gemini[2] but this probably applies to Claude and GPT too, judging from what I read online. Not an endorsement of Gemini's capabilities. I was trying to give AI companies the least amount of money and Gemini subscription had a 2/3 off sale. Not an endorsement of Gemini's capabilities. I was trying to give AI companies the least amount of money and Gemini subscription had a 2/3 off sale. Some of Eliezer's founder effects on the AI alignment/x-safety field, that seem detrimental and persist to
x Wei Dai's Shortform — AI Alignment Forum Wei Dai's Shortform by Wei Dai 1st Mar 2024 1 min read 595 6 This is a special post for quick takes by Wei Dai . Only they can create top-level comments. Comments here also appear on the Quick Takes page and All Posts page . Wei Dai's Shortform 30 Wei Dai 25 Vanessa Kosoy 7 Thomas Kwa 12 Vanessa Kosoy 4 Thomas Kwa 6 Thomas Kwa 5 Jan_Kulveit 1 Raemon 1 Raemon 2 Vladimir_Nesov 25 Wei Dai 17 Wei Dai 9 habryka 4 Wei Dai 4 Vladimir_Nesov 12 Wei Dai 9 Buck 7 Wei Dai 7 Wei Dai 5 habryka 3 Wei Dai 3 habryka 3 Wei Dai 5 Buck 5 Ben Pace 11 Wei Dai 9 Wei Dai 8 W
Explore this link on the map →related reading
- LessWronglesswrong.com
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- AI Pause Will Likely Backfire — EA Forumforum.effectivealtruism.org
- Where I agree and disagree with Eliezer — LessWronglesswrong.com
- Critical review of Christiano's disagreements with Yudkowsky — LessWronglesswrong.com
- Comments on Manheim's "What's in a Pause?" — EA Forumforum.effectivealtruism.org
- I Would Have Solved Alignment, But I Was Worried That Would Advance Timelines — LessWronglesswrong.com
- The Best of LessWrong — LessWronglesswrong.com
- Another (outer) alignment failure story — AI Alignment Forumalignmentforum.org
- The Field of AI Alignment: A Postmortem, and What To Do About It — LessWronglesswrong.com
- Shah and Yudkowsky on alignment failures — LessWronglesswrong.com