Decision theory and dynamic inconsistency – The sideways view
In this post I’ll discuss the last bullet in more detail since I think it’s a bit unusual, it’s not something I’ve written about before, and it’s one fo the main ways my view of decision theory has changed in the last few years. (Note: I think this topic is interesting, and could end up being relevant to the world in some weird-yet-possible situations, but I view it as unrelated to my day job on aligning AI with human interests.) In the transparent version of Newcomb’s problem, you are faced with two transparent boxes (one small and one big). The small box always contains $1,000. The big box contains either $10,000 or $0. You may choose to take the contents of one or both boxes. There is a very accurate predictor, who has placed $10,000 in the big box if and only if they predict that you wouldn’t take the small box regardless of what you see in the big one. Intuitively, once you see the contents of the big box, you really have no reason not to take the small box. For example, if you se
Here is my current take on decision theory: When making a decision after observing X, we should condition (or causally intervene) on statements like “My decision algorithm outputs Y after observing X.” Updating seems like a description of something you do when making good decisions in this way, not part of defining what a good decision is. ( More .) Causal reasoning likewise seems like a description of something you do when making good decisions. Or equivalently: we should use a notion of causality that captures the relationships relevant to decision-making rather than intuitions a
Explore this link on the map →related reading
- Decision theory and dynamic inconsistency — LessWronglesswrong.com
- UDT shows that decision theory is more puzzling than ever — LessWronglesswrong.com
- On Functional Decision Theory :: Wolfgang Schwarzumsu.de
- Updateless Decision Theory — LessWronglesswrong.com
- Thoughts on Updatelessness – The Universe from an Intentional Stancecasparoesterheld.com
- On The Independence Axiom — LessWronglesswrong.com
- UDT shows that decision theory is more puzzling than ever — AI Alignment Forumalignmentforum.org
- Operationalizing FDT — LessWronglesswrong.com
- Updatelessness doesn't solve most problems — LessWronglesswrong.com
- In defense of anthropically updating EDT — LessWronglesswrong.com
- Ingredients of Timeless Decision Theory — LessWronglesswrong.com
- The Shard Theory of Human Valuesturntrout.com