Conspiracy Theories and Where to Find Them on TikTok
Francesco Corso, Francesco Pierri, Gianmarco De Francisci Morales
Abstract
TikTok has skyrocketed in popularity over recent years, especially among younger audiences. However, there are public concerns about the potential of this platform to promote and amplify harmful content. This study presents the first systematic analysis of conspiracy theories on TikTok. By leveraging the official TikTok Research API we collect a longitudinal dataset of 1.5M videos shared in the U.S. over three years. We estimate a lower bound on the prevalence of conspiratorial videos (up to 1000 new videos per month) and evaluate the effects of TikTok's Creativity Program for monetization, observing an overall increase in video duration regardless of content. Lastly, we evaluate the capabilities of state-of-the-art open-weight Large Language Models to identify conspiracy theories from audio transcriptions of videos. While these models achieve high precision in detecting harmful content (up to 96%), their overall performance remains comparable to fine-tuned traditional models such as RoBERTa. Our findings suggest that Large Language Models can serve as an effective tool for supporting content moderation strategies aimed at reducing the spread of harmful content on TikTok.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2d2c0d2b-bbbb-4e32-8d2f-8f20b79fd14cCited by top-tier papers1
Ask how each one uses itBuilds on5
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 11,349 citations
- Robust Speech Recognition via Large-Scale Weak SupervisionAlec Radford, Jong Wook Kim, Tao Xu, Greg Brockman et al.ICML 2023 · 6,966 citations
- Quantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formattingMelanie Sclar, Yejin Choi, Yulia Tsvetkov, Alane SuhrICLR 2024 · 682 citations
- Towards Explainable Harmful Meme Detection through Multimodal Debate between Large Language ModelsHongzhan Lin, Ziyang Luo, Wei Gao, Jing Ma et al.WWW 2024 · 43 citations
- Causal Modeling of Climate Activism on RedditJacopo Lenti, Luca Maria Aiello, Corrado Monti, Gianmarco De Francisci MoralesWWW 2025 · 5 citations
Related papers
- Conspiracy Brokers: Understanding the Monetization of YouTube Conspiracy TheoriesCameron Ballard, Ian Goldstein, Pulak Mehta, Genesis Smothers et al.WWW 2022 · 28 citations
- Among Us: Language of Conspiracy Theorists on Mainstream RedditFrancesco Corso, Giuseppe Russo, Francesco Pierri, Gianmarco De Francisci MoralesACL 2026 · 2 citations
- An Empirical Investigation of Personalization Factors on TikTokMaximilian Boeker, Aleksandra UrmanWWW 2022 · 112 citations
- YouTube Recommendations and Effects on Sharing Across Online Social PlatformsCody Buntain, Richard Bonneau, Jonathan Nagler, Joshua A. TuckerCSCW 2021 · 42 citations
- TikTalk: A Video-Based Dialogue Dataset for Multi-Modal Chitchat in Real WorldHongpeng Lin, Ludan Ruan, Wenke Xia, Peiyu Liu et al.ACM MM 2023 · 10 citations
