Predictive Response Optimization: Using Reinforcement Learning to Fight Online Social Network Abuse
Garrett Wilson, Geoffrey Goh, Yan Jiang, Ajay Gupta, Jiaxuan Wang, David Freeman, Francesco Dinuzzo
摘要
Detecting phishing, spam, fake accounts, data scraping, and other malicious activity in online social networks (OSNs) is a problem that has been studied for well over a decade, with a number of important results. Nearly all existing works on abuse detection have as their goal producing the best possible binary classifier; i.e., one that labels unseen examples as"benign"or"malicious"with high precision and recall. However, no prior published work considers what comes next: what does the service actually do after it detects abuse? In this paper, we argue that detection as described in previous work is not the goal of those who are fighting OSN abuse. Rather, we believe the goal to be selecting actions (e.g., ban the user, block the request, show a CAPTCHA, or"collect more evidence") that optimize a tradeoff between harm caused by abuse and impact on benign users. With this framing, we see that enlarging the set of possible actions allows us to move the Pareto frontier in a way that is unattainable by simply tuning the threshold of a binary classifier. To demonstrate the potential of our approach, we present Predictive Response Optimization (PRO), a system based on reinforcement learning that utilizes available contextual information to predict future abuse and user-experience metrics conditioned on each possible action, and select actions that optimize a multi-dimensional tradeoff between abuse/harm and impact on user experience. We deployed versions of PRO targeted at stopping automated activity on Instagram and Facebook. In both cases our experiments showed that PRO outperforms a baseline classification system, reducing abuse volume by 59% and 4.5% (respectively) with no negative impact to users. We also present several case studies that demonstrate how PRO can quickly and automatically adapt to changes in business constraints, system behavior, and/or adversarial tactics.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Angel or Demon: Investigating the Plasticity Interventions' Impact on Backdoor Threats in Deep Reinforcement LearningOubo Ma, Ruixiao Lin, Yang Dai, Jiahao Chen 等ICML 2026 · 被引用 1 次
- Selling the Dream: How Intimate Insiders and Identity-Based Attackers Disrupt Micro-businessesNazanin Sabri, Arkaprabha Bhattacharya, Sterling Williams-Ceci, Daniel V. Bailey 等USENIX Security 2026
它引用的顶会 Paper8
- Protecting accounts from credential stuffing with password breach alertingKurt Thomas, Jennifer Pullman, Kevin Yeo, Ananth Raghunathan 等USENIX Security 2019 · 被引用 154 次
- DEAR: Deep Reinforcement Learning for Online Advertising Impression in Recommender SystemsXiangyu Zhao, Changsheng Gu, Haoshenglun Zhang, Xiwang Yang 等AAAI 2021 · 被引用 131 次
- Driving 2FA Adoption at Scale: Optimizing Two-Factor Authentication Notification Design PatternsMaximilian Golla, Grant Ho, Marika Lohmus, Monica Pulluri 等USENIX Security 2021 · 被引用 48 次
- Deep Entity Classification: Abusive Account Detection for Online Social NetworksTeng Xu, Gerard Goossen, Huseyin Kerem Cevahir, Sara Khodeir 等USENIX Security 2021 · 被引用 41 次
- Understanding the Behaviors of Toxic Accounts on RedditDeepak Kumar, Jeff T. Hancock, Kurt Thomas, Zakir DurumericWWW 2023 · 被引用 31 次
相关 Paper
- Reinforcement-Learning Based Covert Social Influence OperationsSaurabh Kumar, Valerio La Gatta, Andrea Pugliese, Andrew Pulver 等WWW 2025 · 被引用 3 次
- RABot: Reinforcement-Guided Graph Augmentation for Imbalanced and Noisy Social Bot DetectionLonglong Zhang, Xi Wang, Haotong Du, Yangyi Xu 等AAAI 2026
- Socialbots on Fire: Modeling Adversarial Behaviors of Socialbots via Multi-Agent Hierarchical Reinforcement LearningThai Le, Long Tran-Thanh, Dongwon LeeWWW 2022 · 被引用 10 次
- Exploring and Distilling Multi-Dimensional Clues for Interpretable Social Bot DetectionYi Han, Haiqi Lu, Lizi Liao, Shuhan Zhou 等ACL 2026
- Systemization of Knowledge (SoK): Creating a Research Agenda for Human-Centered Real-Time Risk Detection on Social Media PlatformsAshwaq Alsoubai, Jinkyung Park, Sarvech Qadir, Gianluca Stringhini 等CHI 2024 · 被引用 11 次
