AI Capabilities Can Be Significantly Improved Without Expensive Retraining | Epoch AI
While scaling compute is key to improving LLMs, post-training enhancements can offer gains equivalent to 5-20x more compute at less than 1% of the cost.
AI capabilities can be significantly improved without expensive retraining | Epoch AI The massive computation used to train LLMs and similar foundation models has been one of the main drivers of AI progress in recent years, which has led to the recognition of the “Bitter Lesson”: that general methods that better leverage computational power are ultimately the most effective ( Sutton, 2019 ). The cost of training frontier models has now become so high that only a handful of actors can afford it ( Epoch AI, 2023 ). Our study explores methods of improving performance after training that don’t rel
related reading
- AI progress is about to speed up | Epoch AIepoch.ai
- AI in 2025: gestalt — LessWronglesswrong.com
- As Rocks May Think | Eric Jangevjang.com
- PostTrainBenchposttrainbench.com
- I. From GPT-4 to AGI: Counting the OOMs - SITUATIONAL AWARENESSsituational-awareness.ai
- Composer2.pdfcursor.com
- Pretraining progress is mostly coming from datadwarkesh.com
- The Extreme Inefficiency of RL for Frontier Models - Toby Ordtobyord.com
- Scaling: The State of Play in AIoneusefulthing.org
- My picture of the present in AI — LessWronglesswrong.com
- Human Data is (Probably) More Expensive Than Compute for Training Frontier LLMsddkang.substack.com
- Noam Brown on X: "Implications of Large-Scale Test-Time Compute" / Xx.com