From Apples to Strawberries - by Trevor Chow - Bunnyhopping
This publicly marks the start of the “search” paradigm in modern ML, just as ChatGPT’s launch in 2022 marked the arrival of the “learning” paradigm. With this new paradigm, you should expect progress in ML performance over the next two years to be at least as fast as it was in the last two years. In fact, there are reasons to expect it to be even faster. Here’s why! When I talk about “learning” and “search”, I mean the two “general methods that leverage computation” which Rich Sutton named in The Bitter Lesson. Loosely, “learning” is fitting to patterns in the world, while “search” is finding the best option in a space of possibilities. The Bitter Lesson explains why a single model launch, like o1’s, can catalyse such rapid progress. Since “the most effective” methods in ML are these general methods, progress in ML doesn’t occur steadily, but instead comes in fits and spurts. This is because it relies on researchers finding a technique which gets predictably better with more computing
Three weeks ago, OpenAI released the o1 series. This publicly marks the start of the “search” paradigm in modern ML, just as ChatGPT’s launch in 2022 marked the arrival of the “learning” paradigm. With this new paradigm, you should expect progress in ML performance over the next two years to be at least as fast as it was in the last two years. In fact, there are reasons to expect it to be even faster. Here’s why! When I talk about “learning” and “search”, I mean the two “general methods that leverage computation” which Rich Sutton named in The Bitter Lesson. Loosely, “learning” is…
related reading
- As Rocks May Think | Eric Jangevjang.com
- I. From GPT-4 to AGI: Counting the OOMs - SITUATIONAL AWARENESSsituational-awareness.ai
- Beat GPT-4o at Python by searching with 100 dumb LLaMAsmodal.com
- Spending Inference Time - Kevin Lukevinlu.ai
- AI progress is about to speed up | Epoch AIepoch.ai
- The Scaling Hypothesis · Gwern.netgwern.net
- The Extreme Inefficiency of RL for Frontier Models - Toby Ordtobyord.com
- o3 — LessWronglesswrong.com
- Andrej Karpathy — AGI is still a decade awaydwarkesh.com
- There Are No New Ideas in AI… Only New Datasetsblog.jxmo.io
- The Bitter Lessonincompleteideas.net
- The Bitter Lesson – Yuxi on the Wiredyuxi-liu-wired.github.io