flâneur — a map of the web's best reading

Bayesianism versus conservatism versus Goodhart — AI Alignment Forum

alignmentforum.org · 2,854 words · saved by 1 readers

Key argument: if we use a non-Bayesian conservative approach, such as a minimum over different utility functions, then we better have a good reason a…

x Bayesianism versus conservatism versus Goodhart — AI Alignment Forum Goodhart's Law AI Frontpage 8 Bayesianism versus conservatism versus Goodhart by Stuart_Armstrong 16th Jul 2021 7 min read 2 8 Key argument: if we use a non-Bayesian conservative approach, such as a minimum over different utility functions, then we better have a good reason as to why that would work. But if we have that reason, we can use it to make the whole thing into a Bayesian mix, which can also allow us to trade off that advantage against other possible gains. I've defended using Bayesian averaging of possible utility

Explore this link on the map →

related reading