Algorithmic Improvement Is Probably Faster Than Scaling Now — LessWrong
Back in 2020, a group at OpenAI ran a conceptually simple test to quantify how much AI progress was attributable to algorithmic improvements. They took ImageNet models which were state-of-the-art at various times between 2012 and 2020, and checked how much compute was needed to train each to the level of AlexNet (the state-of-the-art from 2012). Main finding: over ~7 years, the compute required fell by ~44x. In other words, algorithmic progress yielded a compute-equivalent doubling time of ~16 months (though error bars are large in both directions). On the compute side of things, in 2018 a group at OpenAI estimated that the compute spent on the largest training runs was growing exponentially with a doubling rate of ~3.4 months, between 2012 and 2018. So at the time, the rate of improvement from compute scaling was much faster than the rate of improvement from algorithmic progress. (Though algorithmic improvement was still faster than Moore's Law; the compute increases were mostly drive
x Algorithmic Improvement Is Probably Faster Than Scaling Now — LessWrong AI Frontpage 147 Algorithmic Improvement Is Probably Faster Than Scaling Now by johnswentworth 6th Jun 2023 AI Alignment Forum 2 min read 25 147 Ω 49 The Story as of ~4 Years Ago Back in 2020, a group at OpenAI ran a conceptually simple test to quantify how much AI progress was attributable to algorithmic improvements. They took ImageNet models which were state-of-the-art at various times between 2012 and 2020, and checked how much compute was needed to train each to the level of AlexNet (the state-of-the-art from 2012).
related reading
- The least understood driver of AI progress | Epoch AIepoch.ai
- AI progress is about to speed up | Epoch AIepoch.ai
- I. From GPT-4 to AGI: Counting the OOMs - SITUATIONAL AWARENESSsituational-awareness.ai
- Most Algorithmic Progress is Data Progressberen.io
- What's going on with AI progress and trends? (As of 5/2025)redwoodresearch.substack.com
- The nature of LLM algorithmic progress (v2) — LessWronglesswrong.com
- My picture of the present in AI — LessWronglesswrong.com
- AI Timelines — LessWronglesswrong.com
- Will scaling work?dwarkeshpatel.com
- Pretraining progress is mostly coming from datadwarkesh.com
- The Scaling Hypothesis · Gwern.netgwern.net
- [2403.05812] Algorithmic progress in language modelsarxiv.org