Learning to Reason with LLMs | OpenAI
We are introducing OpenAI o1, a new large language model trained with reinforcement learning to perform complex reasoning. o1 thinks before it answers—it can produce a long internal chain of thought before responding to the user. OpenAI o1 ranks in the 89th percentile on competitive programming questions (Codeforces), places among the top 500 students in the US in a qualifier for the USA Math Olympiad (AIME), and exceeds human PhD-level accuracy on a benchmark of physics, biology, and chemistry problems (GPQA). While the work needed to make this new model as easy to use as current models is still ongoing, we are releasing an early version of this model, OpenAI o1-preview, for immediate use in ChatGPT and to trusted API users (opens in a new window) . Our large-scale reinforcement learning algorithm teaches the model how to think productively using its chain of thought in a highly data-efficient training process. We have found that the performance of o1 consistently improves with more r
September 12, 2024 Release Learning to reason with LLMs Contributions Use o1 (opens in a new window) Loading… OpenAI o1 ranks in the 89th percentile on competitive programming questions (Codeforces), places among the top 500 students in the US in a qualifier for the USA Math Olympiad (AIME), and exceeds human PhD-level accuracy on a benchmark of physics, biology, and chemistry problems (GPQA). While the work needed to make this new model as easy to use as current models is still ongoing, we are releasing an early version of this model, OpenAI o1‑preview, for immediate use in ChatGPT and to tru
Explore this link on the map →saved by
related reading
- DeepSeek-R1arxiv.org
- Late Takes on OpenAI o1alexirpan.com
- [2201.11903] Chain of Thought Prompting Elicits Reasoning in Large Language Modelsarxiv.org
- o1 and Reasoning | AndoLogsblog.ando.ai
- The Unintelligibility is Ours: Notes on Chain of Thought1a3orn.com
- Reasoning models | OpenAI APIplatform.openai.com
- GitHub - brexhq/prompt-engineering: Tips and tricks for working with Large Language Models like OpenAI's GPT-4. · GitHubgithub.com
- the-illusion-of-thinking.pdfml-site.cdn-apple.com
- Chain-of-Thought Promptinglearnprompting.org
- GitHub - deepseek-ai/DeepSeek-R1 · GitHubgithub.com
- OpenThoughts3 - A new SOTA Reasoning Data Recipe | OpenThoughtsopenthoughts.ai
- o1: A Technical Primer — LessWronglesswrong.com