HateBuffer: Safeguarding Content Moderators' Mental Well-Being through Hate Speech Content Modification
Subin Park, Jeonghyun Kim, Jeanne Choi, Joseph Seering, Uichin Lee, Sung-Ju Lee
摘要
Hate speech remains a persistent and unresolved challenge in online platforms. Content moderators, working on the front lines to review user-generated content and shield viewers from hate speech, often find themselves unprotected from the mental burden as they continuously engage with offensive language. To safeguard moderators' mental well-being, we designed HateBuffer, which anonymizes targets of hate speech, paraphrases offensive expressions into less offensive forms, and shows the original expressions when moderators opt to see them. Our user study with 80 participants consisted of a simulated hate speech moderation task set on a fictional news platform, followed by semi-structured interviews. Although participants rated the hate severity of comments lower while using HateBuffer, contrary to our expectations, they did not experience improved emotion or reduced fatigue compared with the control group. In interviews, however, participants described HateBuffer as an effective buffer against emotional contagion and the normalization of biased opinions in hate speech. Notably, HateBuffer did not compromise moderation accuracy and even contributed to a slight increase in recall. We explore possible explanations for the discrepancy between the perceived benefits of HateBuffer and its measured impact on mental well-being. We also underscore the promise of text-based content modification techniques as tools for a healthier content moderation environment.
CCS Concepts: • Human-centered computing → Empirical studies in HCI; Empirical studies in collaborative and social computing.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper24
- The Psychological Well-Being of Content Moderators: The Emotional Labor of Commercial Moderation and Avenues for Improving SupportMiriah Steiger, Timir J. Bharucha, Sukrit Venkatagiri, Martin J. Riedl 等CHI 2021 · 被引用 168 次
- The Disagreement Deconvolution: Bringing Machine Learning Performance Metrics In Line With RealityMitchell L. Gordon, Kaitlyn Zhou, Kayur Patel, Tatsunori Hashimoto 等CHI 2021 · 被引用 100 次
- Trauma-Informed Social Media: Towards Solutions for Reducing and Healing Online HarmCarol F. Scott, Gabriela Marcu, Riana Elyse Anderson, Mark W. Newman 等CHI 2023 · 被引用 91 次
- Trans Time: Safety, Privacy, and Content Warnings on a Transgender-Specific Social Media SiteOliver L. Haimson, Justin Buss, Zu Weinger, Denny L. Starks 等CSCW 2020 · 被引用 75 次
- Designing Word Filter Tools for Creator-led Comment ModerationShagun Jhaver, Quan Ze Chen, Detlef Knauss, Amy X. ZhangCHI 2022 · 被引用 70 次
相关 Paper
- Awe Versus Aww: The Effectiveness of Two Kinds of Positive Emotional Stimulation on Stress Reduction for Online Content ModeratorsChristine Linda Cook, Jie Cai, Donghee Yvette WohnCSCW 2022 · 被引用 20 次
- Emotionally Aware Moderation: The Potential of Emotion Monitoring in Shaping Healthier Social Media ConversationsXiaotian Su, Naim Zierau, Soomin Kim, April Yi Wang 等CSCW 2025 · 被引用 3 次
- Beyond Content Exposure: Systemic Factors Driving Moderators' Mental Health Crisis in AfricaNuredin Ali Abdelkadir, Tianling Yang, Shivani Kapania, Kauna Ibrahim Malgwi 等CHI 2026 · 被引用 1 次
- Breaking the Silence: Investigating Which Types of Moderation Reduce Negative Effects of Sexist Social Media ContentJulia Sasse, Jens GrossklagsCSCW 2023 · 被引用 12 次
- Personalizing Content Moderation on Social Media: User Perspectives on Moderation Choices, Interface Design, and LaborShagun Jhaver, Alice Qian Zhang, Quan Ze Chen, Nikhila Natarajan 等CSCW 2023 · 被引用 87 次
