[2502.21321] LLM Post-Training: A Deep Dive into Reasoning Large Language Models
arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website. Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them. Have an idea for a project that will add value for arXiv's community? Learn more about arXivLabs. arXiv Operational Status Get status notifications via email or slack
View PDF HTML (experimental) Abstract:Large Language Models (LLMs) have transformed the natural language processing landscape and brought to life diverse applications. Pretraining on vast web-scale data has laid the foundation for these models, yet the research community is now increasingly shifting focus toward post-training techniques to achieve further breakthroughs. While pretraining provides a broad linguistic foundation, post-training methods enable LLMs to refine their knowledge, improve reasoning, enhance factual accuracy, and align more effectively with user intents and ethical…
saved by
related reading
- [2606.07527] Post-training is (Massive) Supervised Learningarxiv.org
- DeepSeek-R1arxiv.org
- Why Does Reinforcement Learning Generalize? A Feature-Level Mechanistic Study of Post-Training in Large Language Modelsarxiv.org
- Getting Caught Up to Modern LLM Research | Samarth Goeldev.samarthgoel.com
- [2412.06769] Training Large Language Models to Reason in a Continuous Latent Spacearxiv.org
- Training Large Language Models to Reason in a Continuous Latent Spacearxiv.org
- PostTrainBenchposttrainbench.com
- [2602.06337] Can Post-Training Transform LLMs into Causal Reasoners?arxiv.org
- [2602.05910] Chunky Post-Training: Data Driven Failures of Generalizationarxiv.org
- LLM Resourcesforrestbicker.com
- [2502.19402] General Reasoning Requires Learning to Reason from the Get-goar5iv.labs.arxiv.org
- The Ultimate Guide to Fine-Tuning LLMs from Basics to Breakthroughs: An Exhaustive Review of Technologies, Research, Best Practices, Applied Research Challenges and Opportunities (Version 1.0)arxiv.org