DeMod: A Holistic Tool with Explainable Detection and Personalized Modification for Toxicity Censorship
Yaqiong Li, Peng Zhang, Hansu Gu, Tun Lu, Siyuan Qiao, Yubo Shu, Yiyang Shao, Ning Gu
Abstract
Although there have been automated approaches and tools supporting toxicity censorship for social posts, most of them focus on detection. Toxicity censorship is a complex process, wherein detection is just an initial task and a user can have further needs such as rationale understanding and content modification. For this problem, we conduct a need-finding study to investigate people's diverse needs in toxicity censorship and then build a ChatGPT-based censorship tool named DeMod accordingly. DeMod is equipped with the features of explainable De tection and personalized Mod ification, providing fine-grained detection results, detailed explanations, and personalized modification suggestions. We also implemented the tool and recruited 35 Weibo users for evaluation. The results suggest DeMod's multiple strengths like the richness of functionality, the accuracy of censorship, and ease of use. Based on the findings, we further propose several insights into the design of content censorship systems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 919164c4-e514-4738-af07-a8ecfe010f70Cited by top-tier papers4
- Power Echoes: Investigating Moderation Biases in Online Power-Asymmetric ConflictsYaqiong Li, Peng Zhang, Peixu Hou, Kainan Tu et al.CHI 2026 · 2 citations
- Echoes of Norms: Investigating Counterspeech Bots' Influence on Bystanders in Online CommunitiesMengyao Wang, Shuai Ma, Nuo Li, Peng Zhang et al.CHI 2026 · 1 citation
- SparkTales: Facilitating Cross-Language Collaborative Storytelling through Coordinator-AI CollaborationWenxin Zhao, Peng Zhang, Hansu Gu, Haoxuan Zhou et al.CHI 2026 · 1 citation
- CHAIRO: Contextual Hierarchical Analogical Induction and Reasoning Optimization for LLMsHaotian Lu, Yuchen Mou, Bingzhe WuACL 2026
Builds on24
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 11,349 citations
- Don't You Know That You're Toxic: Normalization of Toxicity in Online GamingNicole A. Beres, Julian Frommel, Elizabeth Reid, Regan L. Mandryk et al.CHI 2021 · 235 citations
- Latent Hatred: A Benchmark for Understanding Implicit Hate SpeechMai ElSherief, Caleb Ziems, David Muchlinski, Vaishnavi Anupindi et al.EMNLP 2021 · 159 citations
- Reconsidering Self-Moderation: the Role of Research in Supporting Community-Based Models for Online Content ModerationJoseph SeeringCSCW 2020 · 151 citations
- Human-AI Collaboration via Conditional Delegation: A Case Study of Content ModerationVivian Lai, Samuel Carton, Rajat Bhatnagar, Q. Vera Liao et al.CHI 2022 · 135 citations
Related papers
- Building a Personalized Model for Social Media Textual Content CensorshipBaoxi Liu, Peng Zhang, Yubo Shu, Zhengqing Guan et al.CSCW 2022 · 6 citations
- DetoxLLM: A Framework for Detoxification with ExplanationsMd. Tawkat Islam Khondaker, Muhammad Abdul-Mageed, Laks V. S. LakshmananEMNLP 2024 · 4 citations
- Toxicity Detection is NOT all you Need: Measuring the Gaps to Supporting Volunteer Content Moderators through a User-Centric MethodYang Trista Cao, Lovely-Frances Domingo, Sarah A. Gilbert, Michelle L. Mazurek et al.EMNLP 2024 · 4 citations
- What If Moderation Didn't Mean Suppression? A Case for Personalized Content TransformationRayhan Rashed, Farnaz JahanbakhshCHI 2026 · 1 citation
- "I Cannot Write This Because It Violates Our Content Policy": Understanding Content Moderation Policies and User Experiences in Generative AI ProductsLan Gao, Oscar Chen, Rachel Lee, Nick Feamster et al.USENIX Security 2025
