Too much efficiency makes everything worse: overfitting and the strong version of Goodhart’s law | Jascha’s blog
This blog is intended to be a place to share ideas and results that are too weird, incomplete, or off-topic to turn into an academic paper, but that I think may be important. Let me know what you think! Contact links to the left.
Increased efficiency can sometimes, counterintuitively, lead to worse outcomes. This is true almost everywhere. We will name this phenomenon the strong version of [Goodhart's law](https://en.wikipedia.org/wiki/Goodhart%27s_law). As one example, more efficient centralized tracking of student progress by standardized testing seems like such a good idea that well-intentioned laws [mandate it](https://en.wikipedia.org/wiki/No_Child_Left_Behind_Act). However, testing also incentivizes schools to focus more on teaching students to test well, and less on teaching broadly useful skills. As a result, i
saved by
related reading
- Goodhart's law - Wikipediaen.wikipedia.org
- Goodhart's Law — AI Alignment Forumalignmentforum.org
- Evolution as Backstop for Reinforcement Learning · Gwern.netgwern.net
- Scaling Laws, Carefully | Lil'Loglilianweng.github.io
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- Goodhart's Law Isn't as Useful as You Might Think - Commoncogcommoncog.com
- Goodhart's Law and Why Measurement is Hard — Ribbonfarmribbonfarm.com
- The Best of LessWrong — LessWronglesswrong.com
- Current AIs seem pretty misaligned to meblog.redwoodresearch.org
- Why the tails (sometimes) don’t come apartuniversalprior.substack.com
- Optimality is the tiger, and agents are its teeth — LessWronglesswrong.com
- Meditations On Molochslatestarcodexabridged.com