Utility Maximization = Description Length Minimization — LessWrong
There’s a useful intuitive notion of “optimization” as pushing the world into a small set of states, starting from any of a large number of states. Visually: Yudkowsky and Flint both have notable formalizations of this “optimization as compression” idea. This post presents a formalization of optimization-as-compression grounded in information theory. Specifically: to “optimize” a system is to reduce the number of bits required to represent the system state using a particular encoding. In other words, “optimizing” a system means making it compressible (in the information-theoretic sense) by a particular model. This formalization turns out to be equivalent to expected utility maximization, and allows us to interpret any expected utility maximizer as “trying to make the world look like a particular model”. Before diving into the formalism, we’ll walk through a conceptual example, taken directly from Flint’s Ground of Optimization: building a house. Here’s Flint’s diagram: The key idea her
x Utility Maximization = Description Length Minimization — LessWrong Basic Foundations for Agent Models Information theory Optimization Utility Functions AI Rationality Curated 225 Utility Maximization = Description Length Minimization by johnswentworth 18th Feb 2021 AI Alignment Forum 7 min read 54 225 Ω 72 There’s a useful intuitive notion of “optimization” as pushing the world into a small set of states, starting from any of a large number of states. Visually: Yudkowsky and Flint both have notable formalizations of this “optimization as compression” idea. This post presents a formalization
Explore this link on the map →related reading
- Visual Information Theory -- colah's blogcolah.github.io
- Six (and a half) intuitions for KL divergence — LessWronglesswrong.com
- A Mathematical Theory of Communicationpeople.math.harvard.edu
- Shtetl-Optimized >> Blog Archive >> The First Law of Complexodynamicsscottaaronson.blog
- Optimality is the tiger, and agents are its teeth — LessWronglesswrong.com
- Visual Information Theory -- colah's blogcolah.github.io
- 0812.4360arxiv.org
- Introduction to abstract entropy — LessWronglesswrong.com
- Reward is not the optimization target — LessWronglesswrong.com
- Why The Focus on Expected Utility Maximisers? — LessWronglesswrong.com
- Data Compression Explainedmattmahoney.net
- Pascal's Mugging: Tiny Probabilities of Vast Utilities — LessWronglesswrong.com