Automated Content Moderation Increases Adherence to Community Guidelines
Manoel Horta Ribeiro, Justin Cheng, Robert West
摘要
Online social media platforms use automated moderation systems to remove or reduce the visibility of rule-breaking content. While previous work has documented the importance of manual content moderation, the effects of automated content moderation remain largely unknown. Here, in a large study of Facebook comments (𝑛 = 412M), we used a fuzzy regression discontinuity design to measure the impact of automated content moderation on subsequent rule-breaking behavior (number of comments hidden/deleted) and engagement (number of additional comments posted). We found that comment deletion decreased subsequent rule-breaking behavior in shorter threads (20 or fewer comments), even among other participants, suggesting that the intervention prevented conversations from derailing. Further, the effect of deletion on the affected user's subsequent rule-breaking behavior was longer-lived than its effect on reducing commenting in general, suggesting that users were deterred from rule-breaking but not from commenting. In contrast, hiding (rather than deleting) content had small and statistically insignificant effects. Our results suggest that automated content moderation increases adherence to community guidelines. CCS CONCEPTS • Human-centered computing → Empirical studies in collaborative and social computing.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Bystanders of Online Moderation: Examining the Effects of Witnessing Post-Removal ExplanationsShagun Jhaver, Himanshu Rathi, Koustuv SahaCHI 2024 · 被引用 19 次
- Deplatforming Norm-Violating Influencers on Social Media Reduces Overall Online Attention Toward ThemManoel Horta Ribeiro, Shagun Jhaver, Jordi Cluet-i-Martinell, Marie Reignier-Tayar 等CSCW 2025 · 被引用 8 次
- Filtering Discomforting Recommendations with Large Language ModelsJiahao Liu, Yiyang Shao, Peng Zhang, Dongsheng Li 等WWW 2025 · 被引用 8 次
- Why Should This Article Be Deleted? Transparent Stance Detection in Multilingual Wikipedia Editor DiscussionsLucie-Aimée Kaffee, Arnav Arora, Isabelle AugensteinEMNLP 2023 · 被引用 3 次
- Post Guidance for Online CommunitiesManoel Horta Ribeiro, Robert West, Ryan Lewis, Sanjay KairamCSCW 2025 · 被引用 3 次
它引用的顶会 Paper5
- The Hateful Memes Challenge: Detecting Hate Speech in Multimodal MemesDouwe Kiela, Hamed Firooz, Aravind Mohan, Vedanuj Goswami 等NeurIPS 2020 · 被引用 1,022 次
- SoK: Hate, Harassment, and the Changing Landscape of Online AbuseKurt Thomas, Devdatta Akhawe, Michael D. Bailey, Dan Boneh 等S&P 2021 · 被引用 175 次
- Evaluating the Effectiveness of Deplatforming as a Moderation Strategy on TwitterShagun Jhaver, Christian Boylston, Diyi Yang, Amy S. BruckmanCSCW 2021 · 被引用 168 次
- Do Platform Migrations Compromise Content Moderation? Evidence from r/The_Donald and r/IncelsManoel Horta Ribeiro, Shagun Jhaver, Savvas Zannettou, Jeremy Blackburn 等CSCW 2021 · 被引用 102 次
- Effects of Algorithmic Flagging on Fairness: Quasi-experimental Evidence from WikipediaNathan TeBlunthuis, Benjamin Mako Hill, Aaron HalfakerCSCW 2021 · 被引用 3 次
相关 Paper
- Who should set the Standards? Analysing Censored Arabic Content on Facebook during the Palestine-Israel ConflictWalid Magdy, Hamdy Mubarak, Joni SalminenCHI 2025 · 被引用 9 次
- 'There Has To Be a Lot That We're Missing': Moderating AI-Generated Content on RedditTravis Lloyd, Joseph Reagle, Mor NaamanCSCW 2025 · 被引用 6 次
- Breaking the Silence: Investigating Which Types of Moderation Reduce Negative Effects of Sexist Social Media ContentJulia Sasse, Jens GrossklagsCSCW 2023 · 被引用 12 次
- The Unsung Heroes of Facebook Groups Moderation: A Case Study of Moderation Practices and ToolsTina Kuo, Alicia Hernani, Jens GrossklagsCSCW 2023 · 被引用 25 次
- Content Moderation and the Formation of Online Communities: A Theoretical FrameworkCynthia Dwork, Chris Hays, Jon M. Kleinberg, Manish RaghavanWWW 2024 · 被引用 7 次
