Weight Ensembling Improves Reasoning in Language Models
arxiv.org · 6,400 words · saved by 1 readers
N/A
Published as a conference paper at COLM 2025 Weight Ensembling Improves Reasoning in Language Models Xingyu Dang⋆,1 Christina Baek⋆,2 Kaiyue Wen3 Zico Kolter2 Aditi Raghunathan2 1 Tsinghua University 2 Carnegie Mellon University 3 Stanford University # dangxy20@mails.tsinghua.edu.cn, kbaek@andrew.cmu.edu…
related reading
- DeepSeek-R1arxiv.org
- As Rocks May Think | Eric Jangevjang.com
- State of RL for reasoning LLMs | A. Weersaweers.de
- the-illusion-of-thinking.pdfml-site.cdn-apple.com
- Explore | alphaXivalphaxiv.org
- [2201.11903] Chain of Thought Prompting Elicits Reasoning in Large Language Modelsarxiv.org
- o1 and Reasoning | AndoLogsblog.ando.ai
- Training Large Language Models to Reason in a Continuous Latent Spacearxiv.org
- The State of Reinforcement Learning for LLM Reasoningmagazine.sebastianraschka.com
- [2203.14465] STaR: Bootstrapping Reasoning With Reasoningarxiv.org
- [2504.13837] Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?arxiv.org
- Mode-Conditioning Unlocks Superior Test-Time Scalingarxiv.org