[1907.10323] Fairness in Reinforcement Learning
Decision support systems (e.g., for ecological conservation) and autonomous systems (e.g., adaptive controllers in smart cities) start to be deployed in real applications. Although their operations often impact many users or stakeholders, no fairness consideration is generally taken into account in their design, which could lead to completely unfair outcomes for some users or stakeholders. To tackle this issue, we advocate for the use of social welfare functions that encode fairness and present this general novel problem in the context of (deep) reinforcement learning, although it could possibly be extended to other machine learning tasks.
Decision support systems (e.g., for ecological conservation) and autonomous systems (e.g., adaptive controllers in smart cities) start to be deployed in real applications. Although their operations often impact many users or stakeholders, no fairness consideration is generally taken into account in their design, which could lead to completely unfair outcomes for some users or stakeholders. To tackle this issue, we advocate for the use of social welfare functions that encode fairness and present this general novel problem in the context of (deep) reinforcement learning, although it could possib
Explore this link on the map →related reading
- Papers · Nikhil Garggargnikhil.com
- Why Tool AIs Want to Be Agent AIs · Gwern.netgwern.net
- Frontiers | On Consequentialism and Fairnessfrontiersin.org
- Evolution as Backstop for Reinforcement Learning · Gwern.netgwern.net
- Debugging Reinforcement Learning Systemsandyljones.com
- Relative notions of fairnessfairmlbook.org
- Reward Hacking in Reinforcement Learning | Lil'Loglilianweng.github.io
- [1911.03020] A Human-in-the-loop Framework to Construct Context-aware Mathematical Notions of Outcome Fairnessarxiv.org
- Reward is not the optimization target — LessWronglesswrong.com
- Reward Is Not Enough — LessWronglesswrong.com
- RLHF Bookrlhfbook.com
- Statistical Fairness - Turing Commonsalan-turing-institute.github.io