ModSandbox: Facilitating Online Community Moderation Through Error Prediction and Improvement of Automated Rules
Jean Y. Song, Sangwook Lee, Jisoo Lee, Mina Kim, Juho Kim
摘要
Despite the common use of rule-based tools for online content moderation, human moderators still spend a lot of time monitoring them to ensure they work as intended. Based on surveys and interviews with Reddit moderators who use AutoModerator, we identified the main challenges in reducing false positives and false negatives of automated rules: not being able to estimate the actual effect of a rule in advance and having difficulty figuring out how the rules should be updated. To address these issues, we built ModSandbox, a novel virtual sandbox system that detects possible false positives and false negatives of a rule and visualizes which part of the rule is causing issues. We conducted a comparative, between-subject study with online content moderators to evaluate the effect of ModSandbox in improving automated rules. Results show that ModSandbox can support quickly finding possible false positives and false negatives of automated rules and guide moderators to improve them to reduce future errors.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Linguistically Differentiating Acts and Recalls of Racial Microaggressions on Social MediaUma Sushmitha Gunturi, Anisha Kumar, Xiaohan Ding, Eugenia Ha Rim RhoCSCW 2024 · 被引用 9 次
- Chillbot: Content Moderation in the BackchannelJoseph Seering, Manas Khadka, Nava Haghighi, Tanya Yang 等CSCW 2024 · 被引用 8 次
- End User Authoring of Personalized Content Classifiers: Comparing Example Labeling, Rule Writing, and LLM PromptingLeijie Wang, Kathryn Yurechko, Pranati Dani, Quan Ze Chen 等CHI 2025 · 被引用 7 次
- DeMod: A Holistic Tool with Explainable Detection and Personalized Modification for Toxicity CensorshipYaqiong Li, Peng Zhang, Hansu Gu, Tun Lu 等CSCW 2025 · 被引用 6 次
- MentalImager: Exploring Generative Images for Assisting Support-Seekers' Self-Disclosure in Online Mental Health CommunitiesHan Zhang, Jiaqi Zhang, Yuxiang Zhou, Ryan Louie 等CSCW 2025 · 被引用 4 次
它引用的顶会 Paper5
- Reconsidering Self-Moderation: the Role of Research in Supporting Community-Based Models for Online Content ModerationJoseph SeeringCSCW 2020 · 被引用 151 次
- "At the End of the Day Facebook Does What ItWants": How Users Experience Contesting Algorithmic Content ModerationKristen Vaccaro, Christian Sandvig, Karrie KarahaliosCSCW 2020 · 被引用 79 次
- Designing Word Filter Tools for Creator-led Comment ModerationShagun Jhaver, Quan Ze Chen, Detlef Knauss, Amy X. ZhangCHI 2022 · 被引用 70 次
- Modular Politics: Toward a Governance Layer for Online CommunitiesNathan Schneider, Primavera De Filippi, Seth Frey, Joshua Z. Tan 等CSCW 2021 · 被引用 63 次
- PolicyKit: Building Governance in Online CommunitiesAmy X. Zhang, Grant Hugh, Michael S. BernsteinUIST 2020 · 被引用 60 次
相关 Paper
- Toxicity Detection is NOT all you Need: Measuring the Gaps to Supporting Volunteer Content Moderators through a User-Centric MethodYang Trista Cao, Lovely-Frances Domingo, Sarah A. Gilbert, Michelle L. Mazurek 等EMNLP 2024 · 被引用 4 次
- "Think about it like you're a firefighter": Understanding How Reddit Moderators Use the ModqueueTanvi Bajpai, Eshwar ChandrasekharanCHI 2026 · 被引用 1 次
- Post Guidance for Online CommunitiesManoel Horta Ribeiro, Robert West, Ryan Lewis, Sanjay KairamCSCW 2025 · 被引用 3 次
- "I Cannot Write This Because It Violates Our Content Policy": Understanding Content Moderation Policies and User Experiences in Generative AI ProductsLan Gao, Oscar Chen, Rachel Lee, Nick Feamster 等USENIX Security 2025
- The Unsung Heroes of Facebook Groups Moderation: A Case Study of Moderation Practices and ToolsTina Kuo, Alicia Hernani, Jens GrossklagsCSCW 2023 · 被引用 25 次
