Ideology Prediction from Scarce and Biased Supervision: Learn to Disregard the "What" and Focus on the "How"!
Chen Chen, Dylan Walker, Venkatesh Saligrama
摘要
We propose a novel supervised learning approach for political ideology prediction (PIP) that is capable of predicting out-of-distribution inputs. This problem is motivated by the fact that manual data-labeling is expensive, while self-reported labels are often scarce and exhibit significant selection bias. We propose a novel statistical model that decomposes the document embeddings into a linear superposition of two vectors; a latent neutral context vector independent of ideology, and a latent position vector aligned with ideology. We train an end-to-end model that has intermediate contextual and positional vectors as outputs. At deployment time, our model predicts labels for input documents by exclusively leveraging the predicted positional vectors. On two benchmark datasets we show that our model is capable of outputting predictions even when trained with as little as 5% biased data, and is significantly more accurate than the state-of-the-art. Through crowd-sourcing we validate the neutrality of contextual vectors, and show that context filtering results in ideological concentration, allowing for prediction on out-of-distribution examples.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper1
相关 Paper
- Late Fusion with Triplet Margin Objective for Multimodal Ideology Prediction and AnalysisChangyuan Qiu, Winston Wu, Xinliang Frederick Zhang, Lu WangEMNLP 2022 · 被引用 4 次
- PRISM: A Framework for Producing Interpretable Political Bias Embeddings with Political-Aware Cross-EncoderYiqun Sun, Qiang Huang, Anthony Kum Hoe Tung, Jun YuACL 2025 · 被引用 2 次
- Unsupervised Detection of Contextualized Embedding Bias with Application to IdeologyValentin Hofmann, Janet B. Pierrehumbert, Hinrich SchützeICML 2022 · 被引用 1 次
- An Embedding Model for Estimating Legislative Preferences from the Frequency and Sentiment of TweetsGregory Spell, Brian Guay, Sunshine Hillygus, Lawrence CarinEMNLP 2020 · 被引用 6 次
- Unsupervised Belief Representation Learning with Information-Theoretic Variational Graph Auto-EncodersJinning Li, Huajie Shao, Dachun Sun, Ruijie Wang 等SIGIR 2022 · 被引用 36 次
