Gaming Consensus: Coordinated Manipulation in Crowdsourced Fact-Checking
Nikil Selvam, Jay Baxter, Sophie Hilgard, Brad Miller, Keith Coleman, Ellen Vitercik, Sanmi Koyejo
Abstract
Crowdsourced fact-checking systems have been adopted by major social media companies such as X, Meta, TikTok and Google with the aim of combating misleading information at scale without relying on centralized editorial control. These systems have been developed around a common underlying concept: a bridging mechanism that identifies notes flagging misleading information when they receive support from people with different perspectives rather than simple majority support. To our knowledge the only publicly disclosed bridging algorithms deployed for fact-checking are based on matrix factorization, as deployed by both X and Meta, augmented with additional components addressing abuse, targeted manipulation, and contributor brigades. This work examines the core matrix factorization portion of these systems, presenting theoretical and empirical evaluations of the degree to which coordinated users could vote strategically by leveraging the latent representations to fabricate the appearance of synthetic consensus within the bridging mechanism. Using historic production data, we find that up to 10.7% of lower quality notes could be manipulated above consensus thresholds using less than 10 ratings. We complement these findings with a theoretical analysis, revealing counterintuitively that rating a note as ``Not Helpful'' can increase its helpfulness score, as well as a cost model quantifying manipulation effort. We have developed and deployed mitigations within X's Community Notes algorithm to address synthetic consensus.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 302634fb-e98f-486b-9453-9d7ef03ad7fcBuilds on5
- PoisonRec: An Adaptive Data Poisoning Framework for Attacking Black-box Recommender SystemsJunshuai Song, Zhao Li, Zehong Hu, Yucheng Wu et al.ICDE 2020 · 83 citations
- Supernotes: Driving Consensus in Crowd-Sourced Fact-CheckingSoham De, Michiel A. Bakker, Jay Baxter, Martin SaveskiWWW 2025 · 31 citations
- Beyond the Crowd: LLM-Augmented Community Notes for Governing Health MisinformationJiaying Wu, Zihang Fu, Haonan Wang, Fanxiao Li et al.ACL 2026 · 17 citations
- Exploring and Mitigating Adversarial Manipulation of Voting-Based LeaderboardsYangsibo Huang, Milad Nasr, Anastasios Nikolas Angelopoulos, Nicholas Carlini et al.ICML 2025
- Adversarial Robustness for Tabular Data through Cost and Utility AwarenessKlim Kireev, Bogdan Kulynych, Carmela TroncosoNDSS 2023
Related papers
- Request a Note: How the Request Function Shapes X's Community Notes SystemYuwei Chuai, Shuning Zhang, Ziming Wang, Xin Yi et al.CHI 2026 · 1 citation
- Did the Roll-Out of Community Notes Reduce Engagement With Misinformation on X/Twitter?Yuwei Chuai, Haoye Tian, Nicolas Pröllochs, Gabriele LenziniCSCW 2024 · 62 citations
- Community Fact-Checks Trigger Moral Outrage in Replies to Misleading Posts on Social MediaYuwei Chuai, Anastasia Sergeeva, Gabriele Lenzini, Nicolas PröllochsCHI 2025 · 5 citations
- Will the Crowd Game the Algorithm?: Using Layperson Judgments to Combat Misinformation on Social Media by Downranking Distrusted SourcesZiv Epstein, Gordon Pennycook, David G. RandCHI 2020 · 68 citations
- From TikTok to Telegram: Cross-Platform Efficacy and User Acceptance of Erroneous and Flawless Misinformation InterventionsKatrin Hartwig, Tom Biselli, Franziska Schneider, Immanuel Lamp et al.CHI 2026 · 1 citation
