Are Stereotypes Leading LLMs' Zero-Shot Stance Detection ?
Anthony Dubreuil, Antoine Gourru, Christine Largeron, Amine Trabelsi
Abstract
Large Language Models inherit stereotypes from their pretraining data, leading to biased behavior toward certain social groups in many Natural Language Processing tasks, such as hateful speech detection or sentiment analysis. Surprisingly, the evaluation of this kind of bias in stance detection methods has been largely overlooked by the community. Stance Detection involves labeling a statement as being against, in favor, or neutral towards a specific target and is among the most sensitive NLP tasks, as it often relates to political leanings. In this paper, we focus on the bias of Large Language Models when performing stance detection in a zero-shot setting. We automatically annotate posts in pre-existing stance detection datasets with two attributes: dialect or vernacular of a specific group and text complexity/readability, to investigate whether these attributes influence the model's stance detection decisions. Our results show that LLMs exhibit significant stereotypes in stance detection tasks, such as incorrectly associating pro-marijuana views with low text complexity and African American dialect with opposition to Donald Trump.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 71ea65a6-c6b3-4cf7-9413-97c4b1f01b2eBuilds on2
- From Pretraining Data to Language Models to Downstream Tasks: Tracking the Trails of Political Biases Leading to Unfair NLP ModelsShangbin Feng, Chan Young Park, Yuhan Liu, Yulia TsvetkovACL 2023 · 117 citations
- Fair Text Classification with Wasserstein IndependenceThibaud Leteno, Antoine Gourru, Charlotte Laclau, Rémi Emonet et al.EMNLP 2023 · 3 citations
Related papers
- StereoSet: Measuring stereotypical bias in pretrained language modelsMoin Nadeem, Anna Bethke, Siva ReddyACL 2021
- Large Language Models Discriminate Against Speakers of German DialectsMinh Duc Bui, Carolin Holtermann, Valentin Hofmann, Anne Lauscher et al.EMNLP 2025
- Evaluation of African American Language Bias in Natural Language GenerationNicholas Deas, Jessica Grieser, Shana Kleiner, Desmond Patton et al.EMNLP 2023 · 15 citations
- Probing Toxic Content in Large Pre-Trained Language ModelsNedjma Ousidhoum, Xinran Zhao, Tianqing Fang, Yangqiu Song et al.ACL 2021
- EZ-STANCE: A Large Dataset for English Zero-Shot Stance DetectionChenye Zhao, Cornelia CarageaACL 2024
