[2306.03081] Sequential Monte Carlo Steering of Large Language Models using Probabilistic Programs
arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website. Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them. Have an idea for a project that will add value for arXiv's community? Learn more about arXivLabs. arXiv Operational Status Get status notifications via email or slack
[2306.03081] Sequential Monte Carlo Steering of Large Language Models using Probabilistic Programs Skip to main content arXiv is now an independent nonprofit! Learn more × Search arXiv Press Enter to search · Advanced search --> Computer Science > Artificial Intelligence arXiv:2306.03081 (cs) [Submitted on 5 Jun 2023 ( v1 ), last revised 26 Nov 2023 (this version, v2)] Title: Sequential Monte Carlo Steering of Large Language Models using Probabilistic Programs Authors: Alexander K. Lew , Tan Zhi-Xuan , Gabriel Grand , Vikash K. Mansinghka View a PDF of the paper titled Sequential
Explore this link on the map →related reading
- GenAI Handbookgenai-handbook.github.io
- [2305.18290] Direct Preference Optimization: Your Language Model is Secretly a Reward Modelarxiv.org
- State of RL for reasoning LLMs | A. Weersaweers.de
- Large Language Diffusion Modelsarxiv.org
- Stream of Search (SoS): Learning to Search in Languagearxiv.org
- [2201.11903] Chain of Thought Prompting Elicits Reasoning in Large Language Modelsarxiv.org
- Reasoning as Trajectoriesslhleosun.github.io
- Steering Might Stop Working Soon — LessWronglesswrong.com
- Large Language Model: world models or surface statistics?thegradient.pub
- LLM Resourcesforrestbicker.com
- [2203.02155] Training language models to follow instructions with human feedbackarxiv.org
- Language Models can Solve Computer Tasksarxiv.org