Late Takes on OpenAI o1
I realize how late this is, but I didn’t get a post out while o1 was fresh, and still feel like writing one despite it being cold. (Also, OpenAI just announced they’re going to ship new stuff starting tomorrow so it’s now or never to say something.)
I realize how late this is, but I didn’t get a post out while o1 was fresh, and still feel like writing one despite it being cold. (Also, OpenAI just announced they’re going to ship new stuff starting tomorrow so it’s now or never to say something.) OpenAI o1 is a model release widely believed (but not confirmed) to be a post-trained version of GPT-4o. It is directly trained to spend more time generating and exploring long, internal chains of thought before responding. Why would you do this? Some questions may fundamentally require more partial work or “thinking time” to arrive at the correct
saved by
related reading
- Reflections on OpenAIcalv.info
- My picture of the present in AI — LessWronglesswrong.com
- Learning to reason with LLMs | OpenAIopenai.com
- Reverse engineering OpenAI’s o1interconnects.ai
- AI in 2025: gestalt — LessWronglesswrong.com
- As Rocks May Think | Eric Jangevjang.com
- o1: A Technical Primer — LessWronglesswrong.com
- o3 — LessWronglesswrong.com
- I. From GPT-4 to AGI: Counting the OOMs - SITUATIONAL AWARENESSsituational-awareness.ai
- How AI Is Learning to Think in Secretnickandresen.substack.com
- Things I learned at OpenAI - by Karina Nguyen - sémaphoresemaphore.substack.com
- OpenAI o1 Results on ARC-AGI-Pub | ARC Prizearcprize.org