Understanding and Countering Stereotypes: A Computational Approach to the Stereotype Content Model
Kathleen C. Fraser, Isar Nejadgholi, Svetlana Kiritchenko
摘要
Stereotypical language expresses widely-held beliefs about different social categories. Many stereotypes are overtly negative, while others may appear positive on the surface, but still lead to negative consequences. In this work, we present a computational approach to interpreting stereotypes in text through the Stereotype Content Model (SCM), a comprehensive causal theory from social psychology. The SCM proposes that stereotypes can be understood along two primary dimensions: warmth and competence. We present a method for defining warmth and competence axes in semantic embedding space, and show that the four quadrants defined by this subspace accurately represent the warmth and competence concepts, according to annotated lexicons. We then apply our computational SCM model to textual stereotype data and show that it compares favourably with survey-based studies in the psychological literature. Furthermore, we explore various strategies to counter stereotypical beliefs with anti-stereotypes. It is known that countering stereotypes with antistereotypical examples is one of the most effective ways to reduce biased thinking, yet the problem of generating anti-stereotypes has not been previously studied. Thus, a better understanding of how to generate realistic and effective anti-stereotypes can contribute to addressing pressing societal concerns of stereotyping, prejudice, and discrimination.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Social-Group-Agnostic Bias Mitigation via the Stereotype Content ModelAli Omrani, Alireza Salkhordeh Ziabari, Charles Yu, Preni Golazizian 等ACL 2023 · 被引用 13 次
- Reinforcement Guided Multi-Task Learning Framework for Low-Resource Stereotype DetectionRajkumar Pujari, Erik Oveson, Priyanka Kulkarni, Elnaz NouriACL 2022 · 被引用 11 次
- Discovering Differences in the Representation of People using Contextualized Semantic AxesLi Lucy, Divya Tadimeti, David BammanEMNLP 2022 · 被引用 7 次
- StereoMap: Quantifying the Awareness of Human-like Stereotypes in Large Language ModelsSullam Jeoung, Yubin Ge, Jana DiesnerEMNLP 2023 · 被引用 3 次
- "They are uncultured": Unveiling Covert Harms and Social Threats in LLM Generated ConversationsPreetam Prabhu Srikar Dammu, Hayoung Jung, Anjali Singh, Monojit Choudhury 等EMNLP 2024 · 被引用 3 次
它引用的顶会 Paper6
- Language (Technology) is Power: A Critical Survey of "Bias" in NLPSu Lin Blodgett, Solon Barocas, Hal Daumé III, Hanna M. WallachACL 2020 · 被引用 68 次
- The POLAR Framework: Polar Opposites Enable Interpretability of Pre-Trained Word EmbeddingsBinny Mathew, Sandipan Sikdar, Florian Lemmerich, Markus StrohmaierWWW 2020 · 被引用 40 次
- CrowS-Pairs: A Challenge Dataset for Measuring Social Biases in Masked Language ModelsNikita Nangia, Clara Vania, Rasika Bhalerao, Samuel R. BowmanEMNLP 2020 · 被引用 19 次
- Social Bias Frames: Reasoning about Social and Power Implications of LanguageMaarten Sap, Saadia Gabriel, Lianhui Qin, Dan Jurafsky 等ACL 2020 · 被引用 16 次
- Generating Counter Narratives against Online Hate Speech: Data and StrategiesSerra Sinem Tekiroglu, Yi-Ling Chung, Marco GueriniACL 2020 · 被引用 13 次
相关 Paper
- Words of Warmth: Trust and Sociability Norms for over 26k English WordsSaif M. MohammadACL 2025 · 被引用 5 次
- SCoUT: A Framework for Structured Stereotype Analysis in Language ModelsJinxuan Wu, Bin Li, Xiangyang XueAAAI 2026
- Annotating Dimensions of Social Perception in Text: A Sentence-Level Dataset of Warmth and CompetenceMutaz Ayesh, Saif M. Mohammad, Nedjma OusidhoumACL 2026
- Social Debiasing for Fair Multi-Modal LLMsHarry Cheng, Yangyang Guo, Qing Guo, Ming-Hsuan Yang 等ICCV 2025 · 被引用 1 次
- Stereotyping Norwegian Salmon: An Inventory of Pitfalls in Fairness Benchmark DatasetsSu Lin Blodgett, Gilsinia Lopez, Alexandra Olteanu, Robert Sim 等ACL 2021
