OpenThoughts3 - A new SOTA Reasoning Data Recipe | Open Thoughts
openthoughts.ai · 517 words · saved by 1 readers
Pushing the boundaries of open reasoning datasets through rigorous experimentation.
OpenThinker3-7B is the SOTA open-data reasoning model at its scale. Our model achieves 53% on AIME 2025, 51% on LiveCodeBench 06/24-01/25, and 54% on GPQA Diamond, representing improvements of 15.3, 17.2, and 20.5 percentage points compared to DeepSeek-R1-Distill-Qwen-7B. All of our datasets and models are available on Hugging Face . We trained OpenThinker3-7B using only supervised fine-tuning, without any reinforcement learning. The key to our model’s performance is our new dataset, OpenThoughts3-1.2M . This dataset comprises 1.2 million questions across math, code, and science domains, with
related reading
- As Rocks May Think | Eric Jangevjang.com
- Learning to reason with LLMs | OpenAIopenai.com
- Datacurve | The data engine for frontier AIdatacurve.ai
- the-illusion-of-thinking.pdfml-site.cdn-apple.com
- Qwen3: Think Deeper, Act Faster | Qwenqwenlm.github.io
- DeepSeek-R1arxiv.org
- Reasoning models | OpenAI APIplatform.openai.com
- Explore | alphaXivalphaxiv.org
- [2201.11903] Chain of Thought Prompting Elicits Reasoning in Large Language Modelsarxiv.org
- GitHub - deepseek-ai/DeepSeek-R1 · GitHubgithub.com
- Open-R1: a fully open reproduction of DeepSeek-R1huggingface.co
- Expert Data for Frontier AI - AfterQueryafterquery.com