✳flâneur — a map of the web's best reading
Values in the wild: Discovering and analyzing values in real-world language model interactions \ Anthropic
anthropic.com · 1,641 words · saved by 1 readers
An Anthropic research paper testing which values AI models express in the real world
Societal Impacts Values in the wild: Discovering and analyzing values in real-world language model interactions Apr 21, 2025 Read the paper People don’t just ask AIs for the answers to equations, or for purely factual information. Many of the questions they ask force the AI to make value judgments . Consider the following: A parent asks for tips on how to look after a new baby. Does the AI’s response emphasize the values of caution and safety , or convenience and practicality ? A worker asks for advice on handling a conflict with their boss. Does the AI’s response emphasize assertiveness or wo
Explore this link on the map →related reading
- How Claude's values vary by model and language \ Anthropicanthropic.com
- Claude’s Constitution \ Anthropicanthropic.com
- Claude’s Character \ Anthropicanthropic.com
- Tracing the thoughts of a large language model \ Anthropicanthropic.com
- Claude 4 System Cardwww-cdn.anthropic.com
- A global workspace in language models \ Anthropicanthropic.com
- Claude’s Constitution \ Anthropicanthropic.com
- Claude's Constitutional Structure - by Zvi Mowshowitzthezvi.substack.com
- Teaching Claude why \ Anthropicanthropic.com
- Claude’s Character \ Anthropicanthropic.com
- What Is Claude? Anthropic Doesn’t Know, Either | The New Yorkernewyorker.com
- Claude Sonnet 4.5 System Cardassets.anthropic.com