"I Cannot Write This Because It Violates Our Content Policy": Understanding Content Moderation Policies and User Experiences in Generative AI Products
Lan Gao, Oscar Chen, Rachel Lee, Nick Feamster, Chenhao Tan, Marshini Chetty
摘要
While recent research has focused on developing safeguards for generative AI (GAI) model-level content safety, little is known about how content moderation to prevent malicious content performs for end-users in real-world GAI products. To bridge this gap, we investigated content moderation policies and their enforcement in GAI online tools -consumer-facing web-based GAI applications. We first analyzed content moderation policies of 14 GAI online tools. While these policies are comprehensive in outlining moderation practices, they usually lack details on practical implementations and are not specific about how users can aid in moderation or appeal moderation decisions. Next, we examined user-experienced content moderation successes and failures through Reddit discussions on GAI online tools. We found that although moderation systems succeeded in blocking malicious generations pervasively, users frequently experienced frustration in failures of both moderation systems and user support after moderation. Based on these findings, we suggest improvements for content moderation policy and user experiences in real-world GAI products.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Governance of AI-Generated Content: A Case Study on Social Media PlatformsLan Gao, Abani Ahmed, Oscar Chen, Margaux Reyl 等CHI 2026 · 被引用 3 次
- What Users Ask, Policies Miss: Unveiling the Gap Between Community-Expressed Privacy Concerns and LLM Provider PoliciesZhihuang Liu, Zhen Huang, Ling Hu, Yifan Yang 等USENIX Security 2026
它引用的顶会 Paper26
- Finetuned Language Models are Zero-Shot LearnersJason Wei, Maarten Bosma, Vincent Y. Zhao, Kelvin Guu 等ICLR 2022 · 被引用 4,966 次
- Safe RLHF: Safe Reinforcement Learning from Human FeedbackJosef Dai, Xuehai Pan, Ruiyang Sun, Jiaming Ji 等ICLR 2024 · 被引用 656 次
- Fine-Grained Human Feedback Gives Better Rewards for Language Model TrainingZeqiu Wu, Yushi Hu, Weijia Shi, Nouha Dziri 等NeurIPS 2023 · 被引用 516 次
- Disproportionate Removals and Differing Content Moderation Experiences for Conservative, Transgender, and Black Social Media Users: Marginalization and Moderation Gray AreasOliver L. Haimson, Daniel Delmonaco, Peipei Nie, Andrea WegnerCSCW 2021 · 被引用 287 次
- Synthetic Lies: Understanding AI-Generated Misinformation and Evaluating Algorithmic and Human SolutionsJiawei Zhou, Yixuan Zhang, Qianni Luo, Andrea G. Parker 等CHI 2023 · 被引用 283 次
相关 Paper
- 'There Has To Be a Lot That We're Missing': Moderating AI-Generated Content on RedditTravis Lloyd, Joseph Reagle, Mor NaamanCSCW 2025 · 被引用 6 次
- Exploring the Use of Abusive Generative AI Models on CivitaiYiluo Wei, Yiming Zhu, Pan Hui, Gareth TysonACM MM 2024 · 被引用 11 次
- Understanding User Needs and Attitudes for Privacy Protection Tools in Online Visual Content SharingChun-Wei Chiang, Harry Yizhou Tian, Ming YinCSCW 2025 · 被引用 4 次
- Rule Development in Online Communities Amidst Growth and Generative AIAndy Zhao, Mor Naaman, Lancaster WuCSCW 2026
- "Community Guidelines Make this the Best Party on the Internet": An In-Depth Study of Online Platforms' Content Moderation PoliciesBrennan Schaffner, Arjun Nitin Bhagoji, Siyuan Cheng, Jacqueline Mei 等CHI 2024 · 被引用 25 次
