TAMA: Target-Aware Multilingual Abuse Detection by Cascaded Conditional Multi-Task Learning
Jiyan Liu, Youzheng Liu, Taihang Wang, Yimin Wang, Ye Jiang, Diana Maynard
摘要
Protecting public figures from online abuse requires models that go beyond post-level classification to determine whether abuse is directed at a designated target, characterize the abuse intent, and extract textual evidence. We introduce Target-Aware Multilingual Abuse (TAMA), a benchmark of 9,386 X (Twitter) posts aimed at public figures, with aligned supervision for (i) tri-class target detection, (ii) 12-way fine-grained abuse type classification, and (iii) phrase-level abusive span localization. To exploit the hierarchical coupling of these tasks, we propose Cascaded-MTL, a dependencyaware multi-task framework that conditions downstream predictions on upstream beliefs via three lightweight modules: Cross-Task Feature Fusion (CTF), Task-Adaptive Gating (TAG), and Label-Guided Span Detection (LGSD). Experiments across three multilingual encoders show that Cascaded-MTL consistently yields higher average F1 than single-task and standard multi-task training and delivers robust gains on type classification and span localization. The code and the dataset are released here: https://github. com/zgjiangtoby/CASCADED-MTL. Disclaimer: The examples presented by this paper may be considered offensive or vulgar.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- Unsupervised Cross-lingual Representation Learning at ScaleAlexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary 等ACL 2020 · 被引用 539 次
- DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding SharingPengcheng He, Jianfeng Gao, Weizhu ChenICLR 2023 · 被引用 394 次
- HateDay: Insights from a Global Hate Speech Dataset Representative of a Day on TwitterManuel Tonneau, Diyi Liu, Niyati Malhotra, Scott A. Hale 等ACL 2025 · 被引用 12 次
相关 Paper
- Joint Modelling of Emotion and Abusive Language DetectionSanthosh Rajamanickam, Pushkar Mishra, Helen Yannakoudakis, Ekaterina ShutovaACL 2020 · 被引用 4 次
- KOLD: Korean Offensive Language DatasetYounghoon Jeong, Juhyun Oh, Jongwon Lee, Jaimeen Ahn 等EMNLP 2022 · 被引用 41 次
- Learning From Dictionary: Enhancing Robustness of Machine-Generated Text Detection in Zero-Shot Language via Adversarial TrainingYuanfan Li, Qi Zhou, Zexuan XieICLR 2026
- GenEx: A Commonsense-aware Unified Generative Framework for Explainable Cyberbullying DetectionKrishanu Maity, Raghav Jain, Prince Jha, Sriparna Saha 等EMNLP 2023 · 被引用 4 次
- A Multi-Task Incremental Learning Framework with Category Name Embedding for Aspect-Category Sentiment AnalysisZehui Dai, Cheng Peng, Huajie Chen, Yadong DingEMNLP 2020 · 被引用 29 次
