PolyNarrative: A Multilingual, Multilabel, Multi-domain Dataset for Narrative Extraction from News Articles
Nikolaos Nikolaidis, Nicolas Stefanovitch, Purificação Silvano, Dimitar Iliyanov Dimitrov, Roman Yangarber, Nuno Guimarães, Elisa Sartori, Ion Androutsopoulos, Preslav Nakov, Giovanni Da San Martino, Jakub Piskorski
摘要
We present PolyNarrative, a new multilingual dataset of news articles, annotated for narratives. Narratives are overt or implicit claims, recurring across articles and languages, promoting a specific interpretation or viewpoint on an ongoing topic, often propagating mis/disinformation. We developed two-level taxonomies with coarse-and fine-grained narrative labels for two domains: (i) climate change and (ii) the military conflict between Ukraine and Russia. We collected news articles in four languages (Bulgarian, English, Portuguese, and Russian) related to the two domains and manually annotated them at the paragraph level. We make the dataset publicly available, along with experimental results of several strong baselines that assign narrative labels to news articles at the paragraph or the document level. We believe that this dataset will foster research in narrative detection and enable new research directions towards more multi-domain and highly granular narrative related tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper3
- Unsupervised Cross-lingual Representation Learning at ScaleAlexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary 等ACL 2020 · 被引用 539 次
- Multilingual Multifaceted Understanding of Online News in Terms of Genre, Framing, and Persuasion TechniquesJakub Piskorski, Nicolas Stefanovitch, Nikolaos Nikolaidis, Giovanni Da San Martino 等ACL 2023 · 被引用 14 次
- Disinformation Capabilities of Large Language ModelsIvan Vykopal, Matús Pikuliak, Ivan Srba, Róbert Móro 等ACL 2024
相关 Paper
- KinyaProp: Fine-Grained Propaganda Annotation in KinyarwandaManzi Fabrice Niyigaba, Ivory Yang, Soroush VosoughiACL 2026
- NewsClaims: A New Benchmark for Claim Detection from News with Attribute KnowledgeRevanth Gangi Reddy, Sai Chetan Chinthakindi, Zhenhailong Wang, Yi R. Fung 等EMNLP 2022 · 被引用 13 次
- Conflicts, Villains, Resolutions: Towards models of Narrative Media FramingLea Frermann, Jiatong Li, Shima Khanehzar, Gosia MikolajczakACL 2023 · 被引用 9 次
- Multilingual Previously Fact-Checked Claim RetrievalMatús Pikuliak, Ivan Srba, Róbert Móro, Timo Hromadka 等EMNLP 2023 · 被引用 9 次
- MIPD: Exploring Manipulation and Intention In a Novel Corpus of Polish DisinformationArkadiusz Modzelewski, Giovanni Da San Martino, Pavel Savov, Magdalena Wilczynska 等EMNLP 2024 · 被引用 2 次
