DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
arxiv.org · 6,184 words · saved by 1 readers
N/A
DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models DeepSeek-AI research@deepseek.com…
saved by
related reading
- QwenLong-L1.5: Post-Training Recipe for Long-Context Reasoning and Memory Managementarxiv.org
- DeepSeek-R1arxiv.org
- As Rocks May Think | Eric Jangevjang.com
- Composer2.pdfcursor.com
- Explore | alphaXivalphaxiv.org
- DeepSeek R1's recipe to replicate o1 and the future of reasoning LMsinterconnects.ai
- Inkling: Our Open-Weights Model - Thinking Machines Labthinkingmachines.ai
- Open-R1: a fully open reproduction of DeepSeek-R1huggingface.co
- DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Modelsarxiv.org
- [2505.17373] Value-Guided Search for Efficient Chain-of-Thought Reasoningarxiv.org
- The Illustrated DeepSeek-R1newsletter.languagemodels.co
- The Complete Guide to DeepSeek Models: V3, R1, V4 and Beyondbentoml.com