USENIX Security2025Top-tier venue
"I Cannot Write This Because It Violates Our Content Policy": Understanding Content Moderation Policies and User Experiences in Generative AI Products
Lan Gao, Oscar Chen, Rachel Lee, Nick Feamster, Chenhao Tan, Marshini Chetty
Abstract
While recent research has focused on developing safeguards for generative AI (GAI) model-level content safety, little is known about how content moderation to prevent malicious content performs for end-users in real-world GAI products. To bridge this gap, we investigated content moderation policies and their enforcement in GAI online tools -consumer-facing web-based GAI applications. We first analyzed content moderation policies of 14 GAI online tools. While these policies are comprehensive in outlining moderation practices, they usually lack details on practical implementations and are not specific about how users can aid in moderation or appeal moderation decisions. Next, we examined user-experienced content moderation successes and failures through Reddit discussions on GAI online tools. We found that although moderation systems succeeded in blocking malicious generations pervasively, users frequently experienced frustration in failures of both moderation systems and user support after moderation. Based on these findings, we suggest improvements for content moderation policy and user experiences in real-world GAI products.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d0e3edab-222c-4ea2-ac58-bffad2de76adCited by top-tier papers2
- Governance of AI-Generated Content: A Case Study on Social Media PlatformsLan Gao, Abani Ahmed, Oscar Chen, Margaux Reyl et al.CHI 2026 · 3 citations
- What Users Ask, Policies Miss: Unveiling the Gap Between Community-Expressed Privacy Concerns and LLM Provider PoliciesZhihuang Liu, Zhen Huang, Ling Hu, Yifan Yang et al.USENIX Security 2026
Builds on26
- Finetuned Language Models are Zero-Shot LearnersJason Wei, Maarten Bosma, Vincent Y. Zhao, Kelvin Guu et al.ICLR 2022 · 4,966 citations
- Safe RLHF: Safe Reinforcement Learning from Human FeedbackJosef Dai, Xuehai Pan, Ruiyang Sun, Jiaming Ji et al.ICLR 2024 · 656 citations
- Fine-Grained Human Feedback Gives Better Rewards for Language Model TrainingZeqiu Wu, Yushi Hu, Weijia Shi, Nouha Dziri et al.NeurIPS 2023 · 516 citations
- Disproportionate Removals and Differing Content Moderation Experiences for Conservative, Transgender, and Black Social Media Users: Marginalization and Moderation Gray AreasOliver L. Haimson, Daniel Delmonaco, Peipei Nie, Andrea WegnerCSCW 2021 · 287 citations
- Synthetic Lies: Understanding AI-Generated Misinformation and Evaluating Algorithmic and Human SolutionsJiawei Zhou, Yixuan Zhang, Qianni Luo, Andrea G. Parker et al.CHI 2023 · 283 citations
Related papers
- 'There Has To Be a Lot That We're Missing': Moderating AI-Generated Content on RedditTravis Lloyd, Joseph Reagle, Mor NaamanCSCW 2025 · 6 citations
- Exploring the Use of Abusive Generative AI Models on CivitaiYiluo Wei, Yiming Zhu, Pan Hui, Gareth TysonACM MM 2024 · 11 citations
- Understanding User Needs and Attitudes for Privacy Protection Tools in Online Visual Content SharingChun-Wei Chiang, Harry Yizhou Tian, Ming YinCSCW 2025 · 4 citations
- Rule Development in Online Communities Amidst Growth and Generative AIAndy Zhao, Mor Naaman, Lancaster WuCSCW 2026
- "Community Guidelines Make this the Best Party on the Internet": An In-Depth Study of Online Platforms' Content Moderation PoliciesBrennan Schaffner, Arjun Nitin Bhagoji, Siyuan Cheng, Jacqueline Mei et al.CHI 2024 · 25 citations
