Detecting Propaganda Techniques in Code-Switched Social Media Text
Muhammad Umar Salman, Asif Hanif, Shady Shehata, Preslav Nakov
摘要
Propaganda is a form of communication intended to influence the opinions and the mindset of the public to promote a particular agenda. With the rise of social media, propaganda has spread rapidly, leading to the need for automatic propaganda detection systems. Most work on propaganda detection has focused on high-resource languages, such as English, and little effort has been made to detect propaganda for low-resource languages. Yet, it is common to find a mix of multiple languages in social media communication, a phenomenon known as code-switching. Code-switching combines different languages within the same text, which poses a challenge for automatic systems. Considering this premise, we propose a novel task of detecting propaganda techniques in codeswitched text. To support this task, we create a corpus of 1,030 texts code-switching between English and Roman Urdu, annotated with 20 propaganda techniques at the fragment level. We perform a number of experiments contrasting different experimental setups, and we find that it is important to model the multilinguality directly rather than using translation as well as to use the right fine-tuning strategy. The code and the dataset are publicly available at https://github.com/mbzuai-nlp/ propaganda-codeswitched-text
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper7
- Deberta: decoding-Enhanced Bert with Disentangled AttentionPengcheng He, Xiaodong Liu, Jianfeng Gao, Weizhu ChenICLR 2021 · 被引用 3,729 次
- Unsupervised Cross-lingual Representation Learning at ScaleAlexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary 等ACL 2020 · 被引用 539 次
- DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding SharingPengcheng He, Jianfeng Gao, Weizhu ChenICLR 2023 · 被引用 394 次
- Hate-Speech and Offensive Language Detection in Roman UrduHammad Rizwan, Muhammad Haroon Shakeel, Asim KarimEMNLP 2020 · 被引用 97 次
- Multilingual Multifaceted Understanding of Online News in Terms of Genre, Framing, and Persuasion TechniquesJakub Piskorski, Nicolas Stefanovitch, Nikolaos Nikolaidis, Giovanni Da San Martino 等ACL 2023 · 被引用 14 次
相关 Paper
- Detecting Propaganda Techniques in MemesDimitar Dimitrov, Bishr Bin Ali, Shaden Shaar, Firoj Alam 等ACL 2021
- KinyaProp: Fine-Grained Propaganda Annotation in KinyarwandaManzi Fabrice Niyigaba, Ivory Yang, Soroush VosoughiACL 2026
- Tell me Habibi, is it Real or Fake?Kartik Kuckreja, Parul Gupta, Injy Hamed, Thamar Solorio 等ICLR 2026 · 被引用 10 次
- Detection of Human and Machine-Authored Fake News in UrduMuhammad Zain Ali, Yuxia Wang, Bernhard Pfahringer, Tony C. SmithACL 2025
- From English to Code-Switching: Transfer Learning with Strong Morphological CluesGustavo Aguilar, Thamar SolorioACL 2020 · 被引用 1 次
