flâneur — a map of the web's best reading

o1: A Technical Primer — LessWrong

lesswrong.com · 4,367 words · saved by 1 readers

> TL;DR: In September 2024, OpenAI released o1, its first "reasoning model". This model exhibits remarkable test-time scaling laws, which complete a…

x o1: A Technical Primer — LessWrong Recursive Self-Improvement Scaling Laws Summaries AI Frontpage 175 o1: A Technical Primer by Jesse Hoogland 9th Dec 2024 AI Alignment Forum Linkpost for www.youtube.com 10 min read 19 175 Ω 65 TL;DR : In September 2024, OpenAI released o1, its first "reasoning model". This model exhibits remarkable test-time scaling laws , which complete a missing piece of the Bitter Lesson and open up a new axis for scaling compute. Following Rush and Ritter (2024) and Brown ( 2024a , 2024b ), I explore four hypotheses for how o1 works and discuss some implications for fut

Explore this link on the map →

related reading