✳flâneur — a map of the web's best reading
On how various plans miss the hard bits of the alignment challenge — LessWrong
lesswrong.com · 21,288 words · saved by 1 readers
Nate Soares reviews a dozen plans and proposals for making AI go well. He finds that almost none of them grapple with what he considers the core prob…
x On how various plans miss the hard bits of the alignment challenge — LessWrong 2022 MIRI Alignment Discussion Research Agendas AI Risk Threat Models (AI) AI Curated 322 On how various plans miss the hard bits of the alignment challenge by So8res 12th Jul 2022 AI Alignment Forum 35 min read 91 322 Ω 98 This post has been recorded as part of the LessWrong Curated Podcast, and can be listened to on Spotify , Apple Podcasts , and Libsyn . (As usual, this post was written by Nate Soares with some help and editing from Rob Bensinger.) In my last post , I described a “hard bit” of the challenge of
Explore this link on the map →related reading
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- Current AIs seem pretty misaligned to me — LessWronglesswrong.com
- Shah and Yudkowsky on alignment failures — LessWronglesswrong.com
- Alignment remains a hard, unsolved problem — AI Alignment Forumalignmentforum.org
- A positive case for how we might succeed at prosaic AI alignment — AI Alignment Forumalignmentforum.org
- The Field of AI Alignment: A Postmortem, and What To Do About It — LessWronglesswrong.com
- A Pragmatic Vision for Interpretability — AI Alignment Forumalignmentforum.org
- A central AI alignment problem: capabilities generalization, and the sharp left turn — LessWronglesswrong.com
- Ngo and Yudkowsky on alignment difficulty — LessWronglesswrong.com
- What Is The Alignment Problem? — LessWronglesswrong.com
- (My understanding of) What Everyone in Technical Alignment is Doing and Why — LessWronglesswrong.com