[2410.02683] DailyDilemmas: Revealing Value Preferences of LLMs with Quandaries of Daily Life
arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website. Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them. Have an idea for a project that will add value for arXiv's community? Learn more about arXivLabs. arXiv Operational Status Get status notifications via email or slack
[2410.02683] DailyDilemmas: Revealing Value Preferences of LLMs with Quandaries of Daily Life Skip to main content arXiv is now an independent nonprofit! Learn more × Search arXiv Press Enter to search · Advanced search --> Computer Science > Computation and Language arXiv:2410.02683 (cs) [Submitted on 3 Oct 2024 ( v1 ), last revised 15 Mar 2025 (this version, v3)] Title: DailyDilemmas: Revealing Value Preferences of LLMs with Quandaries of Daily Life Authors: Yu Ying Chiu , Liwei Jiang , Yejin Choi View a PDF of the paper titled DailyDilemmas: Revealing Value Preferences of LLMs
Explore this link on the map →related reading
- Guardian Angels: LLM Personalization for Productivity and Security · Gwern.netgwern.net
- How can LLMs help with value-guided decision making? - AI • Objectives • Instituteai.objectives.institute
- LLM Daydreaming · Gwern.netgwern.net
- [2607.14345] Value Leakage: An LLM's Answers Are Silently Shaped by Its Own Valuesarxiv.org
- [2506.00751] Alignment Revisited: Are Large Language Models Consistent in Stated and Revealed Preferences?arxiv.org
- Where Do LLM Values Come From? — LessWronglesswrong.com
- [2008.02275] Aligning AI With Shared Human Valuesarxiv.org
- [2303.17548] Whose Opinions Do Language Models Reflect?arxiv.org
- Emotion Concepts and their Function in a Large Language Modeltransformer-circuits.pub
- LLM Alignment, ethical and mathematical realism, and the most important actions in davidad's understanding — LessWronglesswrong.com
- Values in the wild: Discovering and analyzing values in real-world language model interactions \ Anthropicanthropic.com
- Discovering Language Model Behaviors with Model-Written Evaluations — LessWronglesswrong.com