Inference-Time Compute Scaling Methods to Improve Reasoning Models
Improving the reasoning abilities of large language models (LLMs) has become one of the hottest topics in 2025, and for good reason. Stronger reasoning skills allow LLMs to tackle more complex problems, making them more capable across a wide range of tasks users care about. In the last few weeks, researchers have shared a large number of new strategies to improve reasoning, including scaling inference-time compute, reinforcement learning, supervised fine-tuning, and distillation. And many approaches combine these techniques for greater effect. This article explores recent research advancements in reasoning-optimized LLMs, with a particular focus on inference-time compute scaling that have emerged since the release of DeepSeek R1. Table of Contents Since most readers are likely already familiar with LLM reasoning models, I will keep the definition short: An LLM-based reasoning model is an LLM designed to solve multi-step problems by generating intermediate steps or structured “thought”
The State of LLM Reasoning Model Inference Inference-Time Compute Scaling Methods to Improve Reasoning Models Sebastian Raschka, PhD Mar 08, 2025 428 11 32 Share Improving the reasoning abilities of large language models (LLMs) has become one of the hottest topics in 2025, and for good reason. Stronger reasoning skills allow LLMs to tackle more complex problems, making them more capable across a wide range of tasks users care about. In the last few weeks, researchers have shared a large number of new strategies to improve reasoning, including scaling inference-time compute, reinforcement learn
Explore this link on the map →related reading
- DeepSeek-R1arxiv.org
- Understanding Reasoning LLMs - by Sebastian Raschka, PhDsebastianraschka.com
- the-illusion-of-thinking.pdfml-site.cdn-apple.com
- Unlocking the Working Memory of Large Language Models for Latent Reasoningarxiv.org
- Explore | alphaXivalphaxiv.org
- LLM Resourcesforrestbicker.com
- The State of Reinforcement Learning for LLM Reasoningmagazine.sebastianraschka.com
- As Rocks May Think | Eric Jangevjang.com
- Scaling in the service of reasoning & model-based ML | Yoshua Bengioyoshuabengio.org
- Generative AI's Act o1: The Reasoning Era Begins | Sequoia Capitalsequoiacap.com
- LLM Inference Performance Engineering: Best Practices | Databricks Blogdatabricks.com
- o1 and Reasoning | AndoLogsblog.ando.ai