Measuring Political Bias in Large Language Models: What Is Said and How It Is Said
Yejin Bang, Delong Chen, Nayeon Lee, Pascale Fung
摘要
We propose to measure political bias in LLMs by analyzing both the content and style of their generated content regarding political issues. Existing benchmarks and measures focus on gender and racial biases. However, political bias exists in LLMs and can lead to polarization and other harms in downstream applications. In order to provide transparency to users, we advocate that there should be fine-grained and explainable measures of political biases generated by LLMs. Our proposed measure looks at different political issues such as reproductive rights and climate change, at both the content (the substance of the generation) and the style (the lexical polarity) of such bias. We measured the political bias in eleven opensourced LLMs and showed that our proposed framework is easily scalable to other topics and is explainable.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- Vision Language Models are BiasedAn Vo, Khai-Nguyen Nguyen, Mohammad Reza Taesiri, Thi Tuong Vy Dang 等ICLR 2026 · 被引用 68 次
- Bias in the Mirror : Are LLMs opinions robust to their own adversarial attacksVirgile Rennard, Christos Xypolopoulos, Michalis VazirgiannisACL 2025 · 被引用 8 次
- Framing Political Bias in Multilingual LLMs Across Pakistani LanguagesAfrozah Nadeem, Mark Dras, Usman NaseemACL 2026 · 被引用 6 次
- LLMS ON TRIAL: Evaluating Judicial Fairness For Large Language ModelsYiran Hu, Zongyue Xue, Haitao Li, Siyuan Zheng 等ICLR 2026 · 被引用 4 次
- Designing Effective AI Explanations for Misinformation Detection: A Comparative Study of Content, Social, and Combined ExplanationsYeaeun Gong, Yifan Liu, Lanyu Shang, Na Wei 等CSCW 2025 · 被引用 2 次
它引用的顶会 Paper13
- Process for Adapting Language Models to Society (PALMS) with Values-Targeted DatasetsIrene Solaiman, Christy DennisonNeurIPS 2021 · 被引用 276 次
- Generative Echo Chamber? Effect of LLM-Powered Search Systems on Diverse Information SeekingNikhil Sharma, Q. Vera Liao, Ziang XiaoCHI 2024 · 被引用 123 次
- From Pretraining Data to Language Models to Downstream Tasks: Tracking the Trails of Political Biases Leading to Unfair NLP ModelsShangbin Feng, Chan Young Park, Yuhan Liu, Yulia TsvetkovACL 2023 · 被引用 117 次
- Language (Technology) is Power: A Critical Survey of "Bias" in NLPSu Lin Blodgett, Solon Barocas, Hal Daumé III, Hanna M. WallachACL 2020 · 被引用 68 次
- "I'm sorry to hear that": Finding New Biases in Language Models with a Holistic Descriptor DatasetEric Michael Smith, Melissa Hall, Melanie Kambadur, Eleonora Presani 等EMNLP 2022 · 被引用 56 次
相关 Paper
- Assessing Reliability and Political Bias In LLMs' Judgements of Formal and Material Inferences With Partisan ConclusionsReto Gubelmann, Ghassen KarrayACL 2025
- Measuring and Mitigating Media Outlet Name Bias in Large Language ModelsSeong-Jin Park, Kang-Min KimEMNLP 2025
- Towards Understanding and Mitigating Social Biases in Language ModelsPaul Pu Liang, Chiyu Wu, Louis-Philippe Morency, Ruslan SalakhutdinovICML 2021 · 被引用 495 次
- Fair or Framed? Political Bias in News Articles Generated by LLMsJunho Yoo, Youhyun ShinEMNLP 2025 · 被引用 1 次
- "Was it "stated" or was it "claimed"?: How linguistic bias affects generative language modelsRoma Patel, Ellie PavlickEMNLP 2021 · 被引用 9 次
