Critical review of Christiano's disagreements with Yudkowsky — LessWrong
This is a review of Paul Christiano's article "where I agree and disagree with Eliezer". Written for the LessWrong 2022 Review. In the existential AI safety community, there is an ongoing debate between positions situated differently on some axis which doesn't have a common agreed-upon name, but where Christiano and Yudkowsky can be regarded as representatives of the two directions[1]. For the sake of this review, I will dub the camps gravitating to the different ends of this axis "Prosers" (after prosaic alignment) and "Poets"[2]. Christiano is a Proser, and so are most people in AI safety groups in the industry. Yudkowsky is a typical Poet [sort-of Poet, but there's an important departure from my characterization below], people in MIRI and the agent foundations community tend to also be such. Prosers tend to be more optimistic, lend more credence to slow takeoff, and place more value on empirical research and solving problems by reproducing them in the lab and iterating on the design
x Critical review of Christiano's disagreements with Yudkowsky — LessWrong AI Takeoff Factored Cognition AI World Modeling Frontpage 176 Critical review of Christiano's disagreements with Yudkowsky by Vanessa Kosoy 27th Dec 2023 AI Alignment Forum 17 min read 40 176 Ω 72 This is a review of Paul Christiano's article " where I agree and disagree with Eliezer ". Written for the LessWrong 2022 Review . In the existential AI safety community, there is an ongoing debate between positions situated differently on some axis which doesn't have a common agreed-upon name, but where Christiano and Yudkows
Explore this link on the map →saved by
related reading
- Where I agree and disagree with Eliezer — LessWronglesswrong.com
- LessWronglesswrong.com
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- Unfalsifiable stories of doom | Mechanize, Inc.mechanize.work
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- A Field Guide to AI Safety—Asteriskasteriskmag.com
- The Best of LessWrong — LessWronglesswrong.com
- A positive case for how we might succeed at prosaic AI alignment — AI Alignment Forumalignmentforum.org
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Shah and Yudkowsky on alignment failures — LessWronglesswrong.com
- The Best of LessWrong — LessWronglesswrong.com