What social attitudes about gender does BERT encode? Leveraging insights from psycholinguistics
Julia Watson, Barend Beekhuizen, Suzanne Stevenson
摘要
Much research has sought to evaluate the degree to which large language models reflect social biases. We complement such work with an approach to elucidating the connections between language model predictions and people's social attitudes. We show how word preferences in a large language model reflect social attitudes about gender, using two datasets from human experiments that found differences in gendered or gender neutral word choices by participants with differing views on gender (progressive, moderate, or conservative). We find that the language model BERT takes into account factors that shape human lexical choice of such language, but may not weigh those factors in the same way people do. Moreover, we show that BERT's predictions most resemble responses from participants with moderate to conservative views on gender. Such findings illuminate how a language model: (1) may differ from people in how it deploys words that signal gender, and (2) may prioritize some social attitudes over others.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper7
- Masked Language Model ScoringJulian Salazar, Davis Liang, Toan Q. Nguyen, Katrin KirchhoffACL 2020 · 被引用 167 次
- Harms of Gender Exclusivity and Challenges in Non-Binary Representation in Language TechnologiesSunipa Dev, Masoud Monajatipoor, Anaelia Ovalle, Arjun Subramonian 等EMNLP 2021 · 被引用 113 次
- Language (Technology) is Power: A Critical Survey of "Bias" in NLPSu Lin Blodgett, Solon Barocas, Hal Daumé III, Hanna M. WallachACL 2020 · 被引用 68 次
- Adhering, Steering, and Queering: Treatment of Gender in Natural Language GenerationYolande A. A. Strengers, Lizhen Qu, Qiongkai Xu, Jarrod KnibbeCHI 2020 · 被引用 31 次
- Toward Gender-Inclusive Coreference ResolutionYang Trista Cao, Hal Daumé IIIACL 2020 · 被引用 20 次
相关 Paper
- Angry Men, Sad Women: Large Language Models Reflect Gendered Stereotypes in Emotion AttributionFlor Miriam Plaza del Arco, Amanda Cercas Curry, Alba Cercas Curry, Gavin Abercrombie 等ACL 2024 · 被引用 8 次
- Using Sociolinguistic Variables to Reveal Changing Attitudes Towards Sexuality and GenderSky CH-Wang, David JurgensEMNLP 2021 · 被引用 6 次
- Evaluating Short-Term Temporal Fluctuations of Social Biases in Social Media Data and Masked Language ModelsYi Zhou, Danushka Bollegala, José Camacho-ColladosEMNLP 2024 · 被引用 3 次
- Comparing human and LLM politeness strategies in free productionHaoran Zhao, Robert D. HawkinsEMNLP 2025 · 被引用 2 次
- A Predictive Factor Analysis of Social Biases and Task-Performance in Pretrained Masked Language ModelsYi Zhou, José Camacho-Collados, Danushka BollegalaEMNLP 2023 · 被引用 1 次
