Teaching LLMs to reason like Bayesians
Google researchers demonstrate how Bayesian teaching through supervised fine-tuning enables LLMs to approximate optimal probabilistic reasoning and generalize to new domains.
Teaching LLMs to reason like Bayesians Skip to main content Teaching LLMs to reason like Bayesians March 4, 2026 Sjoerd van Steenkiste and Tal Linzen, Research Scientists, Google Research We teach LLMs to reason in a Bayesian manner by training them to mimic the predictions of an optimal Bayesian model. Quick links Paper Share Copy link × AI systems based on large language models (LLMs) are increasingly used as agents that interact with users and the world. To do this successfully, LLMs need to construct internal representations of the world and estimate the probability that each of these repr
saved by
related reading
- As Rocks May Think | Eric Jangevjang.com
- DeepSeek-R1arxiv.org
- State of RL for reasoning LLMs | A. Weersaweers.de
- Patterns for Building LLM-based Systems & Productseugeneyan.com
- What We’ve Learned From A Year of Building with LLMs – Applied LLMsapplied-llms.org
- Modifying LLM Beliefs with Synthetic Document Finetuningalignment.anthropic.com
- The State of Reinforcement Learning for LLM Reasoningmagazine.sebastianraschka.com
- Training Large Language Models to Reason in a Continuous Latent Spacearxiv.org
- LLM Post-Training: A Deep Dive into Reasoning Large Language Modelsarxiv.org
- Large Language Models Must Be Taught to Know What They Don't Knowarxiv.org
- [2603.07267] How to Steal Reasoning Without Reasoning Tracesarxiv.org
- [2502.19402] General Reasoning Requires Learning to Reason from the Get-goar5iv.labs.arxiv.org