Coherent extrapolated volition (alignment target) - Arbital
A proposed direction for an extremely well-aligned autonomous superintelligence - do what humans would want, if we knew what the AI knew, thought that fast, and understood ourselves.
x Coherent extrapolated volition (alignment target) — LessWrong Coherent extrapolated volition (alignment target) Edited by Eliezer Yudkowsky , et al. last updated 19th Aug 2023 Introduction "Coherent extrapolated volition" (CEV) is Eliezer Yudkowsky 's proposed thing-to-do with an extremely advanced AGI , if you're extremely confident of your ability to align it on complicated targets. Roughly, a CEV-based superintelligence would do what currently existing humans would want* the AI to do, if counterfactually: We knew everything the AI knew; We could think as fast as the AI and consider all th
Explore this link on the map →related reading
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- ALIGNMENT - by vincent huang - a slice of my mindmindslice.substack.com
- Mirrors and Paintings — LessWronglesswrong.com
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- The hot mess theory of AI misalignment: More intelligent agents behave less coherently | Jascha’s blogsohl-dickstein.github.io
- ⿻ Symbiogenesis vs. Convergent Consequentialism — LessWronglesswrong.com
- Where I agree and disagree with Eliezer — LessWronglesswrong.com
- LLM Alignment, ethical and mathematical realism, and the most important actions in davidad's understanding — LessWronglesswrong.com
- A positive case for how we might succeed at prosaic AI alignment — AI Alignment Forumalignmentforum.org
- The case against AI alignment — LessWronglesswrong.com
- AI Goals Forecast — AI 2027ai-2027.com
- The formal goal is a pointer — LessWronglesswrong.com