flâneur — a map of the web's best reading

Coherent extrapolated volition (alignment target) - Arbital

arbital.com · 6,539 words · saved by 1 readers

A proposed direction for an extremely well-aligned autonomous superintelligence - do what humans would want, if we knew what the AI knew, thought that fast, and understood ourselves.

x Coherent extrapolated volition (alignment target) — LessWrong Coherent extrapolated volition (alignment target) Edited by Eliezer Yudkowsky , et al. last updated 19th Aug 2023 Introduction "Coherent extrapolated volition" (CEV) is Eliezer Yudkowsky 's proposed thing-to-do with an extremely advanced AGI , if you're extremely confident of your ability to align it on complicated targets. Roughly, a CEV-based superintelligence would do what currently existing humans would want* the AI to do, if counterfactually: We knew everything the AI knew; We could think as fast as the AI and consider all th

Explore this link on the map →

related reading