flâneur — a map of the web's best reading

A Case for Model Persona Research — LessWrong

lesswrong.com · 2,987 words · saved by 1 readers

Context: At the Center on Long-Term Risk (CLR) our empirical research agenda focuses on studying (malicious) personas, their relation to generalizati…

x A Case for Model Persona Research — LessWrong LLM Personas AI Frontpage 2025 Top Fifty: 15 % 121 A Case for Model Persona Research by nielsrolf , Maxime Riché , Daniel Tan 15th Dec 2025 5 min read 11 121 Context: At the Center on Long-Term Risk (CLR) our empirical research agenda focuses on studying (malicious) personas, their relation to generalization, and how to prevent misgeneralization, especially given weak overseers (e.g., undetected reward hacking) or underspecified training signals. This has motivated our past research on Emergent Misalignment and Inoculation Prompting , and we want

Explore this link on the map →

related reading