How Do Social Bots Participate in Misinformation Spread? A Comprehensive Dataset and Analysis
Herun Wan, Minnan Luo, Zihan Ma, Guang Dai, Xiang Zhao
Abstract
Social media platforms provide an ideal environment to spread misinformation, where social bots can accelerate the spread. This paper explores the interplay between social bots and misinformation on the Sina Weibo platform. We construct a large-scale dataset that includes annotations for both misinformation and social bots. From the misinformation perspective, the dataset is multimodal, containing 11,393 pieces of misinformation and 16,416 pieces of verified information. From the social bot perspective, this dataset contains 65,749 social bots and 345,886 genuine accounts, annotated using a weakly supervised annotator. Extensive experiments demonstrate the comprehensiveness of the dataset, the clear distinction between misinformation and real information, and the high quality of social bot annotations. Further analysis illustrates that: (i) social bots are deeply involved in information spread; (ii) misinformation with the same topics has similar content, providing the basis of echo chambers, and social bots would amplify this phenomenon; and (iii) social bots generate similar content aiming to manipulate public opinions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d745d9b3-3858-4825-8bb0-e6a66ab82e4cCited by top-tier papers2
- Exploring and Distilling Multi-Dimensional Clues for Interpretable Social Bot DetectionYi Han, Haiqi Lu, Lizi Liao, Shuhan Zhou et al.ACL 2026
- XDAC: XAI-Driven Detection and Attribution of LLM-Generated News Comments in KoreanWooyoung Go, Hyoungshick Kim, Alice Oh, Yongdae KimACL 2025
Builds on34
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Deberta: decoding-Enhanced Bert with Disentangled AttentionPengcheng He, Xiaodong Liu, Jianfeng Gao, Weizhu ChenICLR 2021 · 3,729 citations
- VideoMAE: Masked Autoencoders are Data-Efficient Learners for Self-Supervised Video Pre-TrainingZhan Tong, Yibing Song, Jue Wang, Limin WangNeurIPS 2022 · 2,336 citations
- Swin Transformer V2: Scaling Up Capacity and ResolutionZe Liu, Han Hu, Yutong Lin, Zhuliang Yao et al.CVPR 2022 · 2,138 citations
- GCAN: Graph-aware Co-Attention Networks for Explainable Fake News Detection on Social MediaYi-Ju Lu, Cheng-Te LiACL 2020 · 387 citations
Related papers
- Countering Misinformation via Emotional Response GenerationDaniel Russo, Shane P. Kaszefski-Yaschuk, Jacopo Staiano, Marco GueriniEMNLP 2023 · 4 citations
- Scalable and Generalizable Social Bot Detection through Data SelectionKai-Cheng Yang, Onur Varol, Pik-Mai Hui, Filippo MenczerAAAI 2020 · 385 citations
- SIDA: Social Media Image Deepfake Detection, Localization and Explanation with Large Multimodal ModelZhenglin Huang, Jinwei Hu, Xiangtai Li, Yiwei He et al.CVPR 2025
- MCFEND: A Multi-source Benchmark Dataset for Chinese Fake News DetectionYupeng Li, Haorui He, Jin Bai, Dacheng WenWWW 2024 · 30 citations
- Diffusion of Community Fact-Checked Misinformation on TwitterChiara Patricia Drolsbach, Nicolas PröllochsCSCW 2023 · 47 citations
