Measuring the Persuasiveness of Language Models \ Anthropic
While people have long questioned whether AI models may, at some point, become as persuasive as humans in changing people's minds, there has been limited empirical research into the relationship between model scale and the degree of persuasiveness across model outputs. To address this, we developed a basic method to measure persuasiveness, and used it to compare a variety of Anthropic models across three different generations (Claude 1, 2, and 3), and two classes of models (compact models that are smaller, faster, and more cost-effective, and frontier models that are larger and more capable). Within each class of models (compact and frontier), we find a clear scaling trend across model generations: each successive model generation is rated to be more persuasive than the previous. We also find that our latest and most capable model, Claude 3 Opus, produces arguments that don't statistically differ in their persuasiveness compared to arguments written by humans (Figure 1). We study persu
Societal Impacts Measuring the persuasiveness of language models Apr 9, 2024 While people have long questioned whether AI models may, at some point, become as persuasive as humans in changing people's minds, there has been limited empirical research into the relationship between model scale and the degree of persuasiveness across model outputs. To address this, we developed a basic method to measure persuasiveness, and used it to compare a variety of Anthropic models across three different generations (Claude 1, 2, and 3), and two classes of models (compact models that are smaller, faster, and
Explore this link on the map →related reading
- [2403.14380] On the Conversational Persuasiveness of Large Language Models: A Randomized Controlled Trialarxiv.org
- Discovering Language Model Behaviors with Model-Written Evaluations — LessWronglesswrong.com
- Risks from AI persuasion — AI Alignment Forumalignmentforum.org
- The Dark Forest and Generative AImaggieappleton.com
- Externalized reasoning oversight: a research direction for language model alignment — AI Alignment Forumalignmentforum.org
- BIG-bench/bigbench/benchmark_tasks/convinceme at main · google/BIG-bench · GitHubgithub.com
- Automated Alignment Researchers: Using large language models to scale scalable oversight \ Anthropicanthropic.com
- AI Debaters are More Persuasive when Arguing in Alignment with Their Own Beliefsarxiv.org
- [2303.17548] Whose Opinions Do Language Models Reflect?arxiv.org
- [2306.13000] Apolitical Intelligence? Auditing Delphi's responses on controversial political issues in the USarxiv.org
- Assessing Political Bias in Language Models | Stanford HAIhai.stanford.edu
- Cognitive Biases in Large Language Models — LessWronglesswrong.com