flâneur — a map of the web's best reading

Emotion concepts and their function in a large language model \ Anthropic

anthropic.com · 2,958 words · saved by 8 readers

All modern language models sometimes act like they have emotions. What’s behind these behaviors? Our interpretability team investigates.

Interpretability Emotion concepts and their function in a large language model Apr 2, 2026 Read the paper All modern language models sometimes act like they have emotions. They may say they’re happy to help you, or sorry when they make a mistake. Sometimes they even appear to become frustrated or anxious when struggling with tasks. What’s behind these behaviors? The way modern AI models are trained pushes them to act like a character with human-like characteristics. In addition, these models are known to develop rich and generalizable internal representations of abstract concepts underlying th

Explore this link on the map →

saved by

related reading