Unsupervised Detection of Contextualized Embedding Bias with Application to Ideology
Valentin Hofmann, Janet B. Pierrehumbert, Hinrich Schütze
Abstract
We propose a fully unsupervised method to detect bias in contextualized embeddings. The method leverages the assortative information latently encoded by social networks and combines orthogonality regularization, structured sparsity learning, and graph neural networks to find the embedding subspace capturing this information. As a concrete example, we focus on the phenomenon of ideological bias: we introduce the concept of an ideological subspace, show how it can be found by applying our method to online discussion forums, and present techniques to probe it. Our experiments suggest that the ideological subspace encodes abstract evaluative semantics and reflects changes in the political left-right spectrum during the presidency of Donald Trump.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fa80478c-1cf2-4694-a6fd-5bdc867c3ef6Cited by top-tier papers1
Ask how each one uses itBuilds on6
- Language (Technology) is Power: A Critical Survey of "Bias" in NLPSu Lin Blodgett, Solon Barocas, Hal Daumé III, Hanna M. WallachACL 2020 · 68 citations
- Weakly Supervised Learning of Nuanced Frames for Analyzing Polarization in News MediaShamik Roy, Dan GoldwasserEMNLP 2020 · 42 citations
- Understanding the Language of Political Agreement and Disagreement in Legislative TextsMaryam Davoodi, Eric Waltenburg, Dan GoldwasserACL 2020 · 15 citations
- Do "Undocumented Workers" == "Illegal Aliens"? Differentiating Denotation and Connotation in Vector SpacesAlbert Webson, Zhizhong Chen, Carsten Eickhoff, Ellie PavlickEMNLP 2020 · 8 citations
- We Can Detect Your Bias: Predicting the Political Ideology of News ArticlesRamy Baly, Giovanni Da San Martino, James R. Glass, Preslav NakovEMNLP 2020 · 6 citations
Related papers
- Unsupervised Belief Representation Learning with Information-Theoretic Variational Graph Auto-EncodersJinning Li, Huajie Shao, Dachun Sun, Ruijie Wang et al.SIGIR 2022 · 36 citations
- Understanding Political Polarization via Jointly Modeling Users, Connections and Multimodal Contents on Heterogeneous GraphsHanjia Lyu, Jiebo LuoACM MM 2022 · 12 citations
- PRISM: A Framework for Producing Interpretable Political Bias Embeddings with Political-Aware Cross-EncoderYiqun Sun, Qiang Huang, Anthony Kum Hoe Tung, Jun YuACL 2025 · 2 citations
- An Embedding Model for Estimating Legislative Preferences from the Frequency and Sentiment of TweetsGregory Spell, Brian Guay, Sunshine Hillygus, Lawrence CarinEMNLP 2020 · 6 citations
- Towards Author-informed NLP: Mind the Social BiasInbar Pendzel, Einat MinkovEMNLP 2025
