DiNaM: Disinformation Narrative Mining with Large Language Models
Witold Sosnowski, Arkadiusz Modzelewski, Kinga Skorupska, Adam Wierzbicki
摘要
Disinformation poses a significant threat to democratic societies, public health, and national security. To address this challenge, factchecking experts analyze and track disinformation narratives. However, the process of manually identifying these narratives is highly time-consuming and resource-intensive. In this article, we introduce DiNaM, the first algorithm and structured framework specifically designed for mining disinformation narratives. DiNaM uses a multi-step approach to uncover disinformation narratives. It first leverages Large Language Models (LLMs) to detect false information, then applies clustering techniques to identify underlying disinformation narratives. We evaluated DiNaM's performance using groundtruth disinformation narratives from the EUD-isinfoTest dataset. The evaluation employed the Weighted Chamfer Distance (WCD), which measures the similarity between two sets of embeddings: the ground truth and the predicted disinformation narratives. DiNaM achieved a state-of-the-art WCD score of 0.73, outperforming general-purpose narrative mining methods by a notable margin of 16.4-24.7%. We are releasing DiNaM's codebase and the dataset to the public.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- Narrative Theory for Computational Narrative UnderstandingAndrew Piper, Richard Jean So, David BammanEMNLP 2021 · 被引用 62 次
- Factoring Fact-Checks: Structured Information Extraction from Fact-Checking ArticlesShan Jiang, Simon Baumgartner, Abe Ittycheriah, Cong YuWWW 2020 · 被引用 28 次
- Re-evaluating Word Mover's DistanceRyoma Sato, Makoto Yamada, Hisashi KashimaICML 2022 · 被引用 25 次
- Fighting Fire with Fire: The Dual Role of LLMs in Crafting and Detecting Elusive DisinformationJason Samuel Lucas, Adaku Uchendu, Michiharu Yamashita, Jooyoung Lee 等EMNLP 2023 · 被引用 23 次
- MIPD: Exploring Manipulation and Intention In a Novel Corpus of Polish DisinformationArkadiusz Modzelewski, Giovanni Da San Martino, Pavel Savov, Magdalena Wilczynska 等EMNLP 2024 · 被引用 2 次
相关 Paper
- Disinformation Capabilities of Large Language ModelsIvan Vykopal, Matús Pikuliak, Ivan Srba, Róbert Móro 等ACL 2024
- DiNO: Disinformation Narrative ObserverWitold Sosnowski, Arkadiusz Modzelewski, Kinga Skorupska, Adam WierzbickiACL 2026
- Specious Sites: Tracking the Spread and Sway of Spurious News Stories at ScaleHans W. A. Hanley, Deepak Kumar, Zakir DurumericS&P 2024 · 被引用 18 次
- Tracking the Takes and Trajectories of English-Language News Narratives across Trustworthy and Worrisome WebsitesHans W. A. Hanley, Emily Okabe, Zakir DurumericUSENIX Security 2025
- Evaluation of LLM Vulnerabilities to Being Misused for Personalized Disinformation GenerationAneta Zugecova, Dominik Macko, Ivan Srba, Róbert Móro 等ACL 2025 · 被引用 18 次
