Beyond Semantics: The Unreasonable Effectiveness of Reasonless Intermediate Tokens
This is experimental HTML to improve accessibility. We invite you to report rendering errors. Use Alt+Y to toggle on accessible reporting links and Alt+Shift+Y to toggle off. Learn more about this project and help improve conversions. HTML conversions sometimes display errors due to content that did not convert correctly from the source. This paper uses the following packages that are not yet supported by the HTML conversion tool. Feedback on these issues are not necessary; they are known and are being worked on. Authors: achieve the best HTML results from your LaTeX submissions by following these best practices. Recent impressive results from large reasoning models have been interpreted as a triumph of Chain of Thought (CoT), and especially of the process of training on CoTs sampled from base LLMs in order to help find new reasoning patterns. In this paper, we critically examine that interpretation by investigating how the semantics of intermediate tokens—often anthropomorphized as “
Beyond Semantics: The Unreasonable Effectiveness of Reasonless Intermediate Tokens Kaya Stechly ∗ SCAI, Arizona State University khstechl@asu.edu &Karthik Valmeekam SCAI, Arizona State University kvalmeek@asu.edu &Atharva Gundawar ∗ SCAI, Arizona State University agundawa@asu.edu &Vardhan Palod ∗ SCAI, Arizona State University vpalod@asu.edu &Subbarao Kambhampati SCAI, Arizona State University rao@asu.edu equal contribution Abstract Recent impressive results from large reasoning models have been interpreted as a triumph of Chain of Thought (CoT), and especially of the process of training on Co
Explore this link on the map →related reading
- the-illusion-of-thinking.pdfml-site.cdn-apple.com
- Tracing the thoughts of a large language model \ Anthropicanthropic.com
- DeepSeek-R1arxiv.org
- Explore | alphaXivalphaxiv.org
- Reasoning as Trajectoriesslhleosun.github.io
- Stream of Search (SoS): Learning to Search in Languagearxiv.org
- Reasoning Models Reason Well, Until They Don'tarxiv.org
- The Unintelligibility is Ours: Notes on Chain of Thought1a3orn.com
- [2201.11903] Chain of Thought Prompting Elicits Reasoning in Large Language Modelsarxiv.org
- o1 and Reasoning | AndoLogsblog.ando.ai
- [2510.24941] Can Aha Moments Be Fake? Towards Quantifying Decorative and True Thinking in Chain-of-Thoughtarxiv.org
- Reasoning models don't always say what they think \ Anthropicanthropic.com