The Field of AI Alignment: A Postmortem, and What To Do About It — LessWrong
A policeman sees a drunk man searching for something under a streetlight and asks what the drunk has lost. He says he lost his keys and they both look under the streetlight together. After a few minutes the policeman asks if he is sure he lost them here, and the drunk replies, no, and that he lost them in the park. The policeman asks why he is searching here, and the drunk replies, "this is where the light is". Over the past few years, a major source of my relative optimism on AI has been the hope that the field of alignment would transition from pre-paradigmatic to paradigmatic, and make much more rapid progress. At this point, that hope is basically dead. There has been some degree of paradigm formation, but the memetic competition has mostly been won by streetlighting: the large majority of AI Safety researchers and activists are focused on searching for their metaphorical keys under the streetlight. The memetically-successful strategy in the field is to tackle problems which are ea
x The Field of AI Alignment: A Postmortem, and What To Do About It — LessWrong Best of LessWrong 2024 AI Alignment Fieldbuilding AI Frontpage 334 The Field of AI Alignment: A Postmortem, and What To Do About It by johnswentworth 26th Dec 2024 9 min read 176 334 A policeman sees a drunk man searching for something under a streetlight and asks what the drunk has lost. He says he lost his keys and they both look under the streetlight together. After a few minutes the policeman asks if he is sure he lost them here, and the drunk replies, no, and that he lost them in the park. The policeman asks wh
Explore this link on the map →related reading
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- LessWronglesswrong.com
- PSA: Almost nobody is directly working on superintelligent alignment — LessWronglesswrong.com
- Ngo and Yudkowsky on alignment difficulty — LessWronglesswrong.com
- Alignment remains a hard, unsolved problem — AI Alignment Forumalignmentforum.org
- On how various plans miss the hard bits of the alignment challenge — LessWronglesswrong.com
- Abstract advice to researchers tackling the difficult core problems of AGI alignmenttsvibt.blogspot.com
- Shah and Yudkowsky on alignment failures — LessWronglesswrong.com
- Can we safely automate alignment research? - Joe Carlsmithjoecarlsmith.com
- Focus on the places where you feel shocked everyone's dropping the ball — LessWronglesswrong.com
- (My understanding of) What Everyone in Technical Alignment is Doing and Why — LessWronglesswrong.com