Social Bias Frames
Social Bias Frames is a new way of representing the biases and offensiveness that are implied in language. For example, these frames are meant to distill the implication that "women (candidates) are less qualified" behind the statement "we shouldn’t lower our standards to hire more women." Language has the power to reinforce stereotypes and project social biases onto others, yet most approaches today are limited in how they can detect such biases. For example, the wealth of hate speech or toxic language detection tools today will output yes/no decisions without further explanation, and have been shown to backfire against minority speech. Social Bias Frames makes an important step in distilling potential language biases in a much more holistic way, considering the offensiveness, intent of the speaker, as well as explanations of why the implication is biased, using knowledge about social dynamics and stereotypes. Yes! We collected the Social Bias Inference Corpus (SBIC) which contains 15
Social Bias Frames Social Bias Frames Reasoning about Social and Power Implications of Language Read the paper Watch ACL talk Download the data Data statement MTurk Annotation Template What are Social Bias Frames ? Social Bias Frames is a new way of representing the biases and offensiveness that are implied in language. For example, these frames are meant to distill the implication that "women (candidates) are less qualified" behind the statement "we shouldn’t lower our standards to hire more women." Why did you create Social Bias Frames ? Language has the power to reinforce stereotypes and pr
Explore this link on the map →related reading
- [2306.13000] Apolitical Intelligence? Auditing Delphi's responses on controversial political issues in the USarxiv.org
- Assessing Political Bias in Language Models | Stanford HAIhai.stanford.edu
- [2303.17548] Whose Opinions Do Language Models Reflect?arxiv.org
- The Dark Forest and Generative AImaggieappleton.com
- Cognitive Biases in Large Language Models — LessWronglesswrong.com
- Taboo Your Words — LessWronglesswrong.com
- Discovering Language Model Behaviors with Model-Written Evaluations — LessWronglesswrong.com
- Framing (social sciences) - Wikipediaen.wikipedia.org
- Projects | Karina Nguyenkarinanguyen.com
- GitHub - daviddao/awful-ai: 😈Awful AI is a curated list to track current scary usages of AI - hoping to raise awareness · GitHubgithub.com
- Frame Control — LessWronglesswrong.com
- [2407.12856] AI-AI Bias: large language models favor communications generated by large language modelsarxiv.org