How Generative AI Empowers Attackers and Defenders Across the Trust & Safety Landscape
Patrick Gage Kelley, Steven Rousso-Schindler, Renee Shelby, Kurt Thomas, Allison Woodruff
摘要
Generative AI (GenAI) is a powerful technology poised to reshape Trust & Safety. While misuse by attackers is a growing concern, its defensive capacity remains underexplored. This paper examines these effects through a qualitative study with 43 Trust & Safety experts across five domains: child safety, election integrity, hate and harassment, scams, and violent extremism. Our findings characterize a landscape in which GenAI empowers both attackers and defenders. GenAI dramatically increases the scale and speed of attacks, lowering the barrier to entry for creating harmful content, including sophisticated propaganda and deepfakes. Conversely, defenders envision leveraging GenAI to detect and mitigate harmful content at scale, conduct investigations, deploy persuasive counternarratives, improve moderator wellbeing, and offer user support. This work provides a strategic framework for understanding GenAI's impact on Trust & Safety and charts a path for its responsible use in creating safer online environments.
• Human-centered computing → Human computer interaction (HCI); • Social and professional topics → Computing / technology policy; • Security and privacy → Human and societal aspects of security and privacy; • Computing methodologies → Artificial intelligence.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper26
- Jailbroken: How Does LLM Safety Training Fail?Alexander Wei, Nika Haghtalab, Jacob SteinhardtNeurIPS 2023 · 被引用 2,230 次
- Co-Designing Checklists to Understand Organizational Challenges and Opportunities around Fairness in AIMichael A. Madaio, Luke Stark, Jennifer Wortman Vaughan, Hanna M. WallachCHI 2020 · 被引用 428 次
- Where Responsible AI meets Reality: Practitioner Perspectives on Enablers for Shifting Organizational PracticesBogdana Rakova, Jingying Yang, Henriette Cramer, Rumman ChowdhuryCSCW 2021 · 被引用 326 次
- SoK: Hate, Harassment, and the Changing Landscape of Online AbuseKurt Thomas, Devdatta Akhawe, Michael D. Bailey, Dan Boneh 等S&P 2021 · 被引用 175 次
- Assessing the Fairness of AI Systems: AI Practitioners' Processes, Challenges, and Needs for SupportMichael Madaio, Lisa Egede, Hariharan Subramonyam, Jennifer Wortman Vaughan 等CSCW 2022 · 被引用 149 次
相关 Paper
- Who Gets to Define Safety? A Systematic Review of How Generative AI Research Addresses Youth Online SafetyOzioma Collins Oguine, Adriana Alvarado Garcia, Michael J. Muller, Karla Badillo-UrquiolaCHI 2026 · 被引用 2 次
- How Tech Workers Contend with Hazards of Humanlikeness in Generative AIMark Diaz, Renee Shelby, Eric Corbett, Andrew SmartCHI 2026 · 被引用 1 次
- Families' Vision of Generative AI Agents for Household Safety Against Digital and Physical ThreatsZikai Wen, Lanjing Liu, Yaxing YaoCSCW 2025 · 被引用 3 次
- A Framework to Characterize Reporting on Generative AI UseAgathe Balayn, Varun Nagaraj Rao, Su Lin Blodgett, Aylin Caliskan 等CHI 2026 · 被引用 1 次
- From Symptoms to Systems: An Expert-Guided Approach to Understanding Risks of Generative AI for Eating DisordersAmy A. Winecoff, Kevin KlymanCHI 2026
