Mitigating Bias in Session-based Cyberbullying Detection: A Non-Compromising Approach
Lu Cheng, Ahmadreza Mosallanezhad, Yasin N. Silva, Deborah L. Hall, Huan Liu
摘要
The element of repetition in cyberbullying behavior has directed recent computational studies toward detecting cyberbullying based on a social media session. In contrast to a single text, a session may consist of an initial post and an associated sequence of comments. Yet, emerging efforts to enhance the performance of session-based cyberbullying detection have largely overlooked unintended social biases in existing cyberbullying datasets. For example, a session containing certain demographicidentity terms (e.g., "gay" or "black") is more likely to be classified as an instance of cyberbullying. In this paper, we first show evidence of such bias in models trained on sessions collected from different social media platforms (e.g., Instagram). We then propose a contextaware and model-agnostic debiasing strategy that leverages a reinforcement learning technique, without requiring any extra resources or annotations apart from a pre-defined set of sensitive triggers commonly used for identifying cyberbullying instances. Empirical evaluations show that the proposed strategy can simultaneously alleviate the impacts of the unintended biases and improve the detection performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- A Human-Centered Systematic Literature Review of the Computational Approaches for Online Sexual Risk DetectionAfsaneh Razi, Seunghyun Kim, Ashwaq Alsoubai, Gianluca Stringhini 等CSCW 2021 · 被引用 93 次
- Spanning the Spectrum of Hatred Detection: A Persian Multi-Label Hate Speech Dataset with Annotator RationalesZahra Delbari, Nafise Sadat Moosavi, Mohammad Taher PilehvarAAAI 2024 · 被引用 11 次
- Bias Mitigation for Toxicity Detection via Sequential DecisionsLu Cheng, Ahmadreza Mosallanezhad, Yasin N. Silva, Deborah L. Hall 等SIGIR 2022 · 被引用 8 次
- CoNewsReader: Supporting Comprehensive Understanding and Raising Critical Thoughts on Social Media News Through CommentsKangyu Yuan, Guanzheng Chen, Sizhe Liang, Hehai Lin 等CSCW 2026
它引用的顶会 Paper2
- A Reinforced Generation of Adversarial Examples for Neural Machine TranslationWei Zou, Shujian Huang, Jun Xie, Xinyu Dai 等ACL 2020 · 被引用 66 次
- Demographics Should Not Be the Reason of Toxicity: Mitigating Discrimination in Text Classifications with Instance WeightingGuanhua Zhang, Bing Bai, Junqi Zhang, Kun Bai 等ACL 2020 · 被引用 56 次
相关 Paper
- Improving Cyberbullying Detection with User InteractionSuyu Ge, Lu Cheng, Huan LiuWWW 2021 · 被引用 41 次
- HENIN: Learning Heterogeneous Neural Interaction Networks for Explainable Cyberbullying Detection on Social MediaHsin-Yu Chen, Cheng-Te LiEMNLP 2020 · 被引用 27 次
- Automated Detection of Doxing on TwitterYounes Karimi, Anna Cinzia Squicciarini, Shomir WilsonCSCW 2022 · 被引用 18 次
- GenEx: A Commonsense-aware Unified Generative Framework for Explainable Cyberbullying DetectionKrishanu Maity, Raghav Jain, Prince Jha, Sriparna Saha 等EMNLP 2023 · 被引用 4 次
- Predictive Response Optimization: Using Reinforcement Learning to Fight Online Social Network AbuseGarrett Wilson, Geoffrey Goh, Yan Jiang, Ajay Gupta 等USENIX Security 2025
