[2310.08118] Can Large Language Models Really Improve by Self-critiquing Their Own Plans?
arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website. Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them. Have an idea for a project that will add value for arXiv's community? Learn more about arXivLabs. arXiv Operational Status Get status notifications via email or slack
[2310.08118] Can Large Language Models Really Improve by Self-critiquing Their Own Plans? Skip to main content arXiv is now an independent nonprofit! Learn more × Search arXiv Press Enter to search · Advanced search --> Computer Science > Artificial Intelligence arXiv:2310.08118 (cs) [Submitted on 12 Oct 2023] Title: Can Large Language Models Really Improve by Self-critiquing Their Own Plans? Authors: Karthik Valmeekam , Matthew Marquez , Subbarao Kambhampati View a PDF of the paper titled Can Large Language Models Really Improve by Self-critiquing Their Own Plans?, by Karthik Val
related reading
- Can LLMs Critique and Iterate on Their Own Outputs? | Eric Jangevjang.com
- Subbarao Kambhampati (కంభంపాటి సుబ్బారావు) on X: "So my👇 thread about our papers investigating the verification and self-critiquing inabilities of GPT4 has apparently resonated with a lot of folks. Here is a quick response to several issuetwitter.com
- [2605.20873] PlanningBench: Generating Scalable and Verifiable Planning Data for Evaluating and Training Large Language Modelsarxiv.org
- 2408.11326arxiv.org
- What We’ve Learned From A Year of Building with LLMs – Applied LLMsapplied-llms.org
- Productizing Large Language Modelsblog.replit.com
- The bitter lesson of LLM evalsparsed.com
- Language Models can Solve Computer Tasksarxiv.org
- Scaling Self-Play with Self-Guidancearxiv.org
- Discovering Language Model Behaviors with Model-Written Evaluations — LessWronglesswrong.com
- Position: It's Time to Optimize for Self-Consistencytime-for-consistency.github.io
- LLMs Can Self-Improvearxiv.org