Why Agent Foundations? An Overly Abstract Explanation - LessWrong
Let’s say you’re relatively new to the field of AI alignment. You notice a certain cluster of people in the field who claim that no substantive progress is likely to be made on alignment without firs…
x Why Agent Foundations? An Overly Abstract Explanation — LessWrong Best of LessWrong 2022 Agent Foundations Goodhart's Law AI Curated 321 Why Agent Foundations? An Overly Abstract Explanation by johnswentworth 25th Mar 2022 AI Alignment Forum 9 min read 60 321 Ω 84 Let’s say you’re relatively new to the field of AI alignment. You notice a certain cluster of people in the field who claim that no substantive progress is likely to be made on alignment without first solving various foundational questions of agency. These sound like a bunch of weird pseudophilosophical questions, like “ what does
Explore this link on the map →related reading
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- A positive case for how we might succeed at prosaic AI alignment — AI Alignment Forumalignmentforum.org
- What Is The Alignment Problem? — LessWronglesswrong.com
- Beliefs are Chosen to Serve Goals — LessWronglesswrong.com
- The Plan - 2023 Version — LessWronglesswrong.com
- 2023 letter | Zhengdongzhengdongwang.com
- The Field of AI Alignment: A Postmortem, and What To Do About It — LessWronglesswrong.com
- Shah and Yudkowsky on alignment failures — LessWronglesswrong.com
- On how various plans miss the hard bits of the alignment challenge — LessWronglesswrong.com
- Inner Alignment: Explain like I'm 12 Edition — LessWronglesswrong.com
- Dreams of AI alignment: The danger of suggestive names — LessWronglesswrong.com