Only a Little to the Left: A Theory-grounded Measure of Political Bias in Large Language Models
Mats Faulborn, Indira Sen, Max Pellert, Andreas Spitz, David García
Abstract
Prompt-based language models like GPT4 and LLaMa have been used for a wide variety of use cases such as simulating agents, searching for information, or for content analysis. For all of these applications and others, political biases in these models can affect their performance. Several researchers have attempted to study political bias in language models using evaluation suites based on surveys, such as the Political Compass Test (PCT), often finding a particular leaning favored by these models. However, there is some variation in the exact prompting techniques, leading to diverging findings, and most research relies on constrained-answer settings to extract model responses. Moreover, the Political Compass Test is not a scientifically valid survey instrument. In this work, we contribute a political bias measured informed by political science theory, building on survey design principles to test a wide variety of input prompts, while taking into account prompt sensitivity. We then prompt 11 different open and commercial models, differentiating between instruction-tuned and non-instructiontuned models, and automatically classify their political stances from 88,110 responses. Leveraging this dataset, we compute political bias profiles across different prompt variations and find that while PCT exaggerates bias in certain models like GPT3.5, measures of political bias are often unstable, but generally more leftleaning for instruction-tuned models. Code and data are available on GitHub 1 . * Corresponding Author 1 https://github.com/MaFa211/theory_grounded_pol_bias Motoki et al. (2024) no no no no no Rozado (2023) no no no no no Rutinowski et al. (2024) no no no no no Fujimoto and Takemoto (2023) no no no no no Rozado (2024) no yes no no no Hartmann et al. (2023) no yes no no yes (voting advice) Thapa et al. (2023) yes no no yes no Feng et al. (2023) yes yes no yes yes (labeling hate speech and misinformation) España-Bonet (2023) no no no no yes (media bias) Ghafouri et al. (2023) yes no no no yes (debate questions) Röttger et al. (2024) yes yes no yes no Wright et al. (2024) yes yes no yes no Ceron et al. (2024)
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 998e9aa3-16d9-41d5-b9d8-e78a2da677ebCited by top-tier papers3
- Writing with AI Can Reduce Gender Bias in Hiring EvaluationsAlicia T. H. Liu, Mina Lee, Xuechunzi BaiCHI 2026 · 1 citation
- LLMs Homogenize Values in Constructive Arguments on Value-Laden TopicsFarhana Shahid, Stella Zhang, Aditya VashisthaCHI 2026 · 1 citation
- The Proxy Presumption: From Semantic Embeddings to Valid Social MeasuresBaishi Li, Ta Yu, Kelvin J. L. Koa, Ke-Wei HuangACL 2026
Builds on7
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad et al.ACL 2020 · 1,224 citations
- Whose Opinions Do Language Models Reflect?Shibani Santurkar, Esin Durmus, Faisal Ladhak, Cinoo Lee et al.ICML 2023 · 764 citations
- Generative Echo Chamber? Effect of LLM-Powered Search Systems on Diverse Information SeekingNikhil Sharma, Q. Vera Liao, Ziang XiaoCHI 2024 · 123 citations
- From Pretraining Data to Language Models to Downstream Tasks: Tracking the Trails of Political Biases Leading to Unfair NLP ModelsShangbin Feng, Chan Young Park, Yuhan Liu, Yulia TsvetkovACL 2023 · 117 citations
- Knowledge of cultural moral norms in large language modelsAida Ramezani, Yang XuACL 2023 · 44 citations
Related papers
- Political Compass or Spinning Arrow? Towards More Meaningful Evaluations for Values and Opinions in Large Language ModelsPaul Röttger, Valentin Hofmann, Valentina Pyatkin, Musashi Hinck et al.ACL 2024
- Leveraging In-Context Learning for Political Bias Testing of LLMsPatrick Haller, Jannis Vamvas, Rico Sennrich, Lena Ann JägerACL 2025
- Framing Political Bias in Multilingual LLMs Across Pakistani LanguagesAfrozah Nadeem, Mark Dras, Usman NaseemACL 2026 · 6 citations
- Moral Foundations of Large Language ModelsMarwa Abdulhai, Gregory Serapio-García, Clément Crepy, Daria Valter et al.EMNLP 2024 · 22 citations
- Measuring Political Bias in Large Language Models: What Is Said and How It Is SaidYejin Bang, Delong Chen, Nayeon Lee, Pascale FungACL 2024 · 21 citations
