Too much efficiency makes everything worse: overfitting and the strong version of Goodhart’s law | Jascha’s blog
This blog is intended to be a place to share ideas and results that are too weird, incomplete, or off-topic to turn into an academic paper, but that I think may be important. Let me know what you think! Contact links to the left.
Increased efficiency can sometimes, counterintuitively, lead to worse outcomes. This is true almost everywhere. We will name this phenomenon the strong version of [Goodhart's law](https://en.wikipedia.org/wiki/Goodhart%27s_law). As one example, more efficient centralized tracking of student progress by standardized testing seems like such a good idea that well-intentioned laws [mandate it](https://en.wikipedia.org/wiki/No_Child_Left_Behind_Act). However, testing also incentivizes schools to focus more on teaching students to test well, and less on teaching broadly useful skills. As a result, i
Explore this link on the map →saved by
related reading
- Goodhart's law - Wikipediaen.wikipedia.org
- Goodhart's Law — AI Alignment Forumalignmentforum.org
- The Best of LessWrong — LessWronglesswrong.com
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- Goodhart's Law Isn't as Useful as You Might Think - Commoncogcommoncog.com
- Goodhart's Law and Why Measurement is Hard — Ribbonfarmribbonfarm.com
- Current AIs seem pretty misaligned to meblog.redwoodresearch.org
- Evolution as Backstop for Reinforcement Learning · Gwern.netgwern.net
- Why the tails (sometimes) don’t come apartuniversalprior.substack.com
- Optimality is the tiger, and agents are its teeth — LessWronglesswrong.com
- Overfitting in Machine Learning: What It Is and How to Prevent Itelitedatascience.com
- Meditations On Molochslatestarcodexabridged.com