How can LLMs help with value-guided decision making? — AI • Objectives • Institute
AOI is excited to be presenting at a NeurIPS workshop focused on morality in human psychology and AI! Our recent work draws on human moral cognition and the latest research on moral representation in large language models (LLMs). For this project, we asked, how could we use LLMs to represent value for real-world action decision making? Read the full paper here Imagine your robotic AI assistant is shopping for you in a grocery store and its goal is to buy your favorite snacks. Why shouldn’t it steal all the snacks, pushing people out of the way as it gets to them? Humans seek not just to maximize a selfish reward, but also to cooperate with others, have a positive impact on the world, and follow societal norms. Aligned AI must likewise take into account not just how much we care about snacks, but also honesty, cooperation, and (not) stealing. Modern AI agents largely follow the highly successful reinforcement learning framework to make decisions for how to act. They choose the action wi
How can LLMs help with value-guided decision making? Dec 11 Written By Colleen McKenzie Image generated with DALL-E Anna Leshinskaya AOI is excited to be presenting at a NeurIPS workshop focused on morality in human psychology and AI! Our recent work draws on human moral cognition and the latest research on moral representation in large language models (LLMs). For this project, we asked, how could we use LLMs to represent value for real-world action decision making? Read the full paper here Imagine your robotic AI assistant is shopping for you in a grocery store and its goal is to buy your fav
Explore this link on the map →related reading
- LLM Powered Autonomous Agents | Lil'Loglilianweng.github.io
- The Persona Selection Model: Why AI Assistants might Behave like Humansalignment.anthropic.com
- Guardian Angels: LLM Personalization for Productivity and Security · Gwern.netgwern.net
- [2410.02683] DailyDilemmas: Revealing Value Preferences of LLMs with Quandaries of Daily Lifearxiv.org
- The state of LLM ethical decision-making - Imbueimbue.com
- LLM Alignment, ethical and mathematical realism, and the most important actions in davidad's understanding — LessWronglesswrong.com
- Alignment Is Proven To Be Solvable - by SE Gygesverysane.ai
- [2008.02275] Aligning AI With Shared Human Valuesarxiv.org
- [2405.17345] Exploring and steering the moral compass of Large Language Modelsarxiv.org
- Where Do LLM Values Come From? — LessWronglesswrong.com
- Emotion Concepts and their Function in a Large Language Modeltransformer-circuits.pub
- The Computational Anatomy of Human Values — LessWronglesswrong.com