flâneur — a map of the web's best reading

Values in the wild: Discovering and analyzing values in real-world language model interactions \ Anthropic

anthropic.com · 1,641 words · saved by 1 readers

An Anthropic research paper testing which values AI models express in the real world

Societal Impacts Values in the wild: Discovering and analyzing values in real-world language model interactions Apr 21, 2025 Read the paper People don’t just ask AIs for the answers to equations, or for purely factual information. Many of the questions they ask force the AI to make value judgments . Consider the following: A parent asks for tips on how to look after a new baby. Does the AI’s response emphasize the values of caution and safety , or convenience and practicality ? A worker asks for advice on handling a conflict with their boss. Does the AI’s response emphasize assertiveness or wo

Explore this link on the map →

related reading