Against ubiquitous alignment taxes - LessWrong
It is often argued that any alignment technique that works primarily by constraining the capabilities of an AI system to be within some bounds cannot work because it imposes too high an 'alignment ta…
x Against ubiquitous alignment taxes — LessWrong AI Risk Alignment Tax AI Frontpage 59 Against ubiquitous alignment taxes by beren 6th Mar 2023 2 min read 10 59 Crossposted from my personal blog . It is often argued that any alignment technique that works primarily by constraining the capabilities of an AI system to be within some bounds cannot work because it imposes too high an 'alignment tax' on the ML system. The argument is that people will either refuse to apply any method that has an alignment tax, or else they will be outcompeted by those who do. I think that this argument is applied t
Explore this link on the map →saved by
related reading
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- Plans A, B, C, and D for misalignment risk — LessWronglesswrong.com
- Current AIs seem pretty misaligned to me — LessWronglesswrong.com
- Off Target | CNAScnas.org
- A positive case for how we might succeed at prosaic AI alignment — AI Alignment Forumalignmentforum.org
- Questions about the Future of AI - by Dwarkesh Pateldwarkesh.com
- Another (outer) alignment failure story — AI Alignment Forumalignmentforum.org
- AI Pause Will Likely Backfire — EA Forumforum.effectivealtruism.org
- AI Alignment Cannot Be Top-Down | AI Frontiersai-frontiers.org
- Alignment can be the “Military-Grade Engineering” of AI | AE Studioae.studio
- Can we safely automate alignment research? - Joe Carlsmithjoecarlsmith.com