Multilingual estimation of political-party positioning: From label aggregation to long-input Transformers
Dmitry Nikolaev, Tanise Ceron, Sebastian Padó
Abstract
Scaling analysis is a technique in computational political science that assigns a political actor (e.g. politician or party) a score on a predefined scale based on a (typically long) body of text (e.g. a parliamentary speech or an election manifesto). For example, political scientists have often used the left–right scale to systematically analyse political landscapes of different countries. NLP methods for automatic scaling analysis can find broad application provided they (i) are able to deal with long texts and (ii) work robustly across domains and languages. In this work, we implement and compare two approaches to automatic scaling analysis of political-party manifestos: label aggregation, a pipeline strategy relying on annotations of individual statements from the manifestos, and long-input-Transformer-based models, which compute scaling values directly from raw text. We carry out the analysis of the Comparative Manifestos Project dataset across 41 countries and 27 languages and find that the task can be efficiently solved by state-of-the-art models, with label aggregation producing the best results.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c7d6c7f4-a61e-4ddf-a25e-0c61c706bfddBuilds on4
- Big Bird: Transformers for Longer SequencesManzil Zaheer, Guru Guruganesh, Kumar Avinava Dubey, Joshua Ainslie et al.NeurIPS 2020 · 3,159 citations
- MPNet: Masked and Permuted Pre-training for Language UnderstandingKaitao Song, Xu Tan, Tao Qin, Jianfeng Lu et al.NeurIPS 2020 · 1,957 citations
- Long Range Arena : A Benchmark for Efficient TransformersYi Tay, Mostafa Dehghani, Samira Abnar, Yikang Shen et al.ICLR 2021 · 881 citations
- Text-Based Ideal PointsKeyon Vafa, Suresh Naidu, David M. BleiACL 2020 · 37 citations
Related papers
- Measuring scalar constructs in social science with LLMsHauke Licht, Rupak Sarkar, Patrick Y. Wu, Pranav Goel et al.EMNLP 2025
- SCALE: Towards Collaborative Content Analysis in Social Science with Large Language Model Agents and Human InterventionChengshuai Zhao, Zhen Tan, Chau-Wai Wong, Xinyan Zhao et al.ACL 2025 · 8 citations
- Measuring Political Bias in Large Language Models: What Is Said and How It Is SaidYejin Bang, Delong Chen, Nayeon Lee, Pascale FungACL 2024 · 21 citations
- Understanding Politics via Contextualized Discourse ProcessingRajkumar Pujari, Dan GoldwasserEMNLP 2021 · 6 citations
- Ruddit: Norms of Offensiveness for English Reddit CommentsRishav Hada, Sohi Sudhir, Pushkar Mishra, Helen Yannakoudakis et al.ACL 2021
