DeMod: A Holistic Tool with Explainable Detection and Personalized Modification for Toxicity Censorship
Yaqiong Li, Peng Zhang, Hansu Gu, Tun Lu, Siyuan Qiao, Yubo Shu, Yiyang Shao, Ning Gu
摘要
Although there have been automated approaches and tools supporting toxicity censorship for social posts, most of them focus on detection. Toxicity censorship is a complex process, wherein detection is just an initial task and a user can have further needs such as rationale understanding and content modification. For this problem, we conduct a need-finding study to investigate people's diverse needs in toxicity censorship and then build a ChatGPT-based censorship tool named DeMod accordingly. DeMod is equipped with the features of explainable De tection and personalized Mod ification, providing fine-grained detection results, detailed explanations, and personalized modification suggestions. We also implemented the tool and recruited 35 Weibo users for evaluation. The results suggest DeMod's multiple strengths like the richness of functionality, the accuracy of censorship, and ease of use. Based on the findings, we further propose several insights into the design of content censorship systems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Power Echoes: Investigating Moderation Biases in Online Power-Asymmetric ConflictsYaqiong Li, Peng Zhang, Peixu Hou, Kainan Tu 等CHI 2026 · 被引用 2 次
- Echoes of Norms: Investigating Counterspeech Bots' Influence on Bystanders in Online CommunitiesMengyao Wang, Shuai Ma, Nuo Li, Peng Zhang 等CHI 2026 · 被引用 1 次
- SparkTales: Facilitating Cross-Language Collaborative Storytelling through Coordinator-AI CollaborationWenxin Zhao, Peng Zhang, Hansu Gu, Haoxuan Zhou 等CHI 2026 · 被引用 1 次
- CHAIRO: Contextual Hierarchical Analogical Induction and Reasoning Optimization for LLMsHaotian Lu, Yuchen Mou, Bingzhe WuACL 2026
它引用的顶会 Paper24
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 被引用 11,349 次
- Don't You Know That You're Toxic: Normalization of Toxicity in Online GamingNicole A. Beres, Julian Frommel, Elizabeth Reid, Regan L. Mandryk 等CHI 2021 · 被引用 235 次
- Latent Hatred: A Benchmark for Understanding Implicit Hate SpeechMai ElSherief, Caleb Ziems, David Muchlinski, Vaishnavi Anupindi 等EMNLP 2021 · 被引用 159 次
- Reconsidering Self-Moderation: the Role of Research in Supporting Community-Based Models for Online Content ModerationJoseph SeeringCSCW 2020 · 被引用 151 次
- Human-AI Collaboration via Conditional Delegation: A Case Study of Content ModerationVivian Lai, Samuel Carton, Rajat Bhatnagar, Q. Vera Liao 等CHI 2022 · 被引用 135 次
相关 Paper
- Building a Personalized Model for Social Media Textual Content CensorshipBaoxi Liu, Peng Zhang, Yubo Shu, Zhengqing Guan 等CSCW 2022 · 被引用 6 次
- DetoxLLM: A Framework for Detoxification with ExplanationsMd. Tawkat Islam Khondaker, Muhammad Abdul-Mageed, Laks V. S. LakshmananEMNLP 2024 · 被引用 4 次
- Toxicity Detection is NOT all you Need: Measuring the Gaps to Supporting Volunteer Content Moderators through a User-Centric MethodYang Trista Cao, Lovely-Frances Domingo, Sarah A. Gilbert, Michelle L. Mazurek 等EMNLP 2024 · 被引用 4 次
- What If Moderation Didn't Mean Suppression? A Case for Personalized Content TransformationRayhan Rashed, Farnaz JahanbakhshCHI 2026 · 被引用 1 次
- "I Cannot Write This Because It Violates Our Content Policy": Understanding Content Moderation Policies and User Experiences in Generative AI ProductsLan Gao, Oscar Chen, Rachel Lee, Nick Feamster 等USENIX Security 2025
