Lambretta: Learning to Rank for Twitter Soft Moderation
Pujan Paudel, Jeremy Blackburn, Emiliano De Cristofaro, Savvas Zannettou, Gianluca Stringhini
Abstract
To curb the problem of false information, social media platforms like Twitter started adding warning labels to content discussing debunked narratives, with the goal of providing more context to their audiences. Unfortunately, these labels are not applied uniformly and leave large amounts of false content unmoderated. This paper presents LAMBRETTA, a system that automatically identifies tweets that are candidates for soft moderation using Learning To Rank (LTR). We run Lambretta on Twitter data to moderate false claims related to the 2020 US Election and find that it flags over 20 times more tweets than Twitter, with only 3.93% false positives and 18.81% false negatives, outperforming alternative state-of-the-art methods based on keyword extraction and semantic search. Overall, LAMBRETTA assists human moderators in identifying and flagging false information on social media.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ebf55b47-1d2e-446e-9160-285e58adf70dCited by top-tier papers9
- Moderating New Waves of Online Hate with Chain-of-Thought Reasoning in Large Language ModelsNishant Vishwamitra, Keyan Guo, Farhan Tajwar Romit, Isabelle Ondracek et al.S&P 2024 · 29 citations
- Specious Sites: Tracking the Spread and Sway of Spurious News Stories at ScaleHans W. A. Hanley, Deepak Kumar, Zakir DurumericS&P 2024 · 18 citations
- Deciphering Textual Authenticity: A Generalized Strategy through the Lens of Large Language Semantics for Detecting Human vs. Machine-Generated TextMazal Bethany, Brandon Wherry, Emet Bethany, Nishant Vishwamitra et al.USENIX Security 2024 · 13 citations
- PIXELMOD: Improving Soft Moderation of Visual Misleading Information on TwitterPujan Paudel, Chen Ling, Jeremy Blackburn, Gianluca StringhiniUSENIX Security 2024 · 4 citations
- Enabling Contextual Soft Moderation on Social Media through Contrastive Textual DeviationPujan Paudel, Mohammad Hammas Saeed, Rebecca Auger, Chris Wells et al.USENIX Security 2024 · 3 citations
Builds on12
- MPNet: Masked and Permuted Pre-training for Language UnderstandingKaitao Song, Xu Tan, Tao Qin, Jianfeng Lu et al.NeurIPS 2020 · 1,957 citations
- GCAN: Graph-aware Co-Attention Networks for Explainable Fake News Detection on Social MediaYi-Ju Lu, Cheng-Te LiACL 2020 · 387 citations
- Detecting Credential Spearphishing in Enterprise SettingsGrant Ho, Aashish Sharma, Mobin Javed, Vern Paxson et al.USENIX Security 2017 · 94 citations
- "Go eat a bat, Chang!": On the Emergence of Sinophobic Behavior on Web Communities in the Face of COVID-19Fatemeh Tahmasbi, Leonard Schild, Chen Ling, Jeremy Blackburn et al.WWW 2021 · 92 citations
- Adapting Security Warnings to Counter Online DisinformationBen Kaiser, Jerry Wei, Eli Lucherini, Kevin Lee et al.USENIX Security 2021 · 81 citations
Related papers
- The Impact of Twitter Labels on Misinformation Spread and User Engagement: Lessons from Trump's Election TweetsOrestis Papakyriakopoulos, Ellen P. GoodmannWWW 2022 · 53 citations
- That is a Known Lie: Detecting Previously Fact-Checked ClaimsShaden Shaar, Nikolay Babulkov, Giovanni Da San Martino, Preslav NakovACL 2020 · 26 citations
- Cross-modal Ambiguity Learning for Multimodal Fake News DetectionYixuan Chen, Dongsheng Li, Peng Zhang, Jie Sui et al.WWW 2022 · 325 citations
- Attacking Misinformation Detection Using Adversarial Examples Generated by Language ModelsPiotr Przybyla, Euan McGill, Horacio SaggionEMNLP 2025 · 1 citation
- Multilingual Detection of Personal Employment Status on TwitterManuel Tonneau, Dhaval Adjodah, João Palotti, Nir Grinberg et al.ACL 2022 · 16 citations
