Reformative Hypocrisy, and Paying Close Enough Attention to Selectively Reward It. — LessWrong
People often attack frontier AI labs for "hypocrisy" when the labs admit publicly that AI is an extinction threat to humanity. Often these attacks i…
x Reformative Hypocrisy, and Paying Close Enough Attention to Selectively Reward It. — LessWrong AI Frontpage 53 Reformative Hypocrisy, and Paying Close Enough Attention to Selectively Reward It. by Andrew_Critch 11th Sep 2024 3 min read 13 53 People often attack frontier AI labs for "hypocrisy" when the labs admit publicly that AI is an extinction threat to humanity. Often these attacks ignore the difference between various kinds of hypocrisy, some of which are good, including what I'll call "reformative hypocrisy". Attacking good kinds of hypocrisy can be actively harmful for humanity's abil
Explore this link on the map →saved by
related reading
- LessWronglesswrong.com
- Current AIs seem pretty misaligned to me — LessWronglesswrong.com
- The Possessed Machines: Dostoevsky's Demons and the Coming AGI Catastrophepossessedmachines.com
- Current AIs seem pretty misaligned to meblog.redwoodresearch.org
- We're already in AI takeoff — LessWronglesswrong.com
- The Best of LessWrong — LessWronglesswrong.com
- A positive case for how we might succeed at prosaic AI alignment — AI Alignment Forumalignmentforum.org
- Moltbook: After The First Weekend - by Scott Alexanderastralcodexten.com
- Your AIs don't do what you want. This is really badrewardhacking.org
- A guide to the AI tribes - by Michel Justen - What is thismicheljusten.substack.com
- 1a3orn's Shortform — LessWronglesswrong.com
- The Magnitude of His Own Folly — LessWronglesswrong.com