Simplistic Collection and Labeling Practices Limit the Utility of Benchmark Datasets for Twitter Bot Detection
Chris Hays, Zachary Schutzman, Manish Raghavan, Erin Walk, Philipp Zimmer
摘要
Accurate bot detection is necessary for the safety and integrity of online platforms. It is also crucial for research on the influence of bots in elections, the spread of misinformation, and financial market manipulation. Platforms deploy infrastructure to flag or remove automated accounts, but their tools and data are not publicly available. Thus, the public must rely on third-party bot detection. These tools employ machine learning and often achieve near-perfect performance for classification on existing datasets, suggesting bot detection is accurate, reliable and fit for use in downstream applications. We provide evidence that this is not the case and show that high performance is attributable to limitations in dataset collection and labeling rather than sophistication of the tools. Specifically, we show that simple decision rules — shallow decision trees trained on a small number of features — achieve near-state-of-the-art performance on most available datasets and that bot detection datasets, even when combined together, do not generalize well to out-of-sample datasets. Our findings reveal that predictions are highly dependent on each dataset’s collection and labeling procedures rather than fundamental differences between bots and humans. These results have important implications for both transparency in sampling and labeling procedures and potential biases in research using existing bot detection tools for pre-processing.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- BotMoE: Twitter Bot Detection with Community-Aware Mixtures of Modal-Specific ExpertsYuhan Liu, Zhaoxuan Tan, Heng Wang, Shangbin Feng 等SIGIR 2023 · 被引用 54 次
- Human-centered NLP Fact-checking: Co-Designing with Fact-checkers using Matchmaking for AIHoujiang Liu, Anubrata Das, Alexander Boltz, Didi Zhou 等CSCW 2024 · 被引用 23 次
- Bots, Elections, and Controversies: Twitter Insights from Brazil's Polarised ElectionsDiogo PachecoWWW 2024 · 被引用 13 次
- Identifying Risky Vendors in Cryptocurrency P2P MarketplacesTaro Tsuchiya, Alejandro Cuevas Villalba, Nicolas ChristinWWW 2024 · 被引用 9 次
- How Do Social Bots Participate in Misinformation Spread? A Comprehensive Dataset and AnalysisHerun Wan, Minnan Luo, Zihan Ma, Guang Dai 等EMNLP 2025 · 被引用 3 次
它引用的顶会 Paper3
- Scalable and Generalizable Social Bot Detection through Data SelectionKai-Cheng Yang, Onur Varol, Pik-Mai Hui, Filippo MenczerAAAI 2020 · 被引用 385 次
- Heterogeneity-Aware Twitter Bot Detection with Relational Graph TransformersShangbin Feng, Zhaoxuan Tan, Rui Li, Minnan LuoAAAI 2022 · 被引用 138 次
- Disagree? You Must be a Bot! How Beliefs Shape Twitter Profile PerceptionsMagdalena Wischnewski, Rebecca Bernemann, Thao Ngo, Nicole C. KrämerCHI 2021 · 被引用 22 次
相关 Paper
- SoK: Machine Learning for Misinformation DetectionMadelyne Xiao, Jonathan R. MayerUSENIX Security 2025
- BotBR: Social Bot Detection with Balanced Feature Fusion and Reliability-Enhanced Graph LearningQilong Lin, Jingya ZhouSIGIR 2025 · 被引用 3 次
- "Better Be Computer or I'm Dumb": A Large-Scale Evaluation of Humans as Audio Deepfake DetectorsKevin Warren, Tyler Tucker, Anna Crowder, Daniel Olszewski 等CCS 2024 · 被引用 9 次
- BIC: Twitter Bot Detection with Text-Graph Interaction and Semantic ConsistencyZhenyu Lei, Herun Wan, Wenqian Zhang, Shangbin Feng 等ACL 2023 · 被引用 28 次
- Bot Meets Shortcut: How Can LLMs Aid in Handling Unknown Invariance OOD Scenarios?Shiyan Zheng, Herun Wan, Minnan Luo, Junhang HuangAAAI 2026
