Multimodal Categorization of Crisis Events in Social Media
Mahdi Abavisani, Liwei Wu, Shengli Hu, Joel R. Tetreault, Alejandro Jaimes
摘要
Recent developments in image classification and natural language processing, coupled with the rapid growth in social media usage, have enabled fundamental advances in detecting breaking events around the world in real-time. Emergency response is one such area that stands to gain from these advances. By processing billions of texts and images a minute, events can be automatically detected to enable emergency response workers to better assess rapidly evolving situations and deploy resources accordingly. To date, most event detection techniques in this area have focused on image-only or text-only approaches, limiting detection performance and impacting the quality of information delivered to crisis response teams. In this paper, we present a new multimodal fusion method that leverages both images and texts as input. In particular, we introduce a cross-attention module that can filter uninformative and misleading components from weak modalities on a sample by sample basis. In addition, we employ a multimodal graph-based approach to stochastically transition between embeddings of different multimodal pairs during training to better regularize the learning process as well as dealing with limited training data by constructing new matched pairs from different samples. We show that our method outperforms the unimodal approaches and strong multimodal baselines by a large margin on three crisis-related tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Graph Neural Architecture Search Under Distribution ShiftsYijian Qin, Xin Wang, Ziwei Zhang, Pengtao Xie 等ICML 2022 · 被引用 41 次
- Expanding Large Pre-trained Unimodal Models with Multimodal Information Injection for Image-Text Multimodal ClassificationTao Liang, Guosheng Lin, Mingyang Wan, Tianrui Li 等CVPR 2022 · 被引用 39 次
- Aligning Vision to Language: Annotation-Free Multimodal Knowledge Graph Construction for Enhanced LLMs ReasoningJunming Liu, Siyuan Meng, Yanting Gao, Song Mao 等ICCV 2025 · 被引用 34 次
- Multimodal Graph Neural Architecture Search under Distribution ShiftsJie Cai, Xin Wang, Haoyang Li, Ziwei Zhang 等AAAI 2024 · 被引用 20 次
- Natural Disaster Tweets Classification Using Multimodal DataMohammad Basit, Bashir Alam, Zubaida Fatima, Salman ShaikhEMNLP 2023 · 被引用 12 次
它引用的顶会 Paper3
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel 等ICLR 2020 · 被引用 7,418 次
- VL-BERT: Pre-training of Generic Visual-Linguistic RepresentationsWeijie Su, Xizhou Zhu, Yue Cao, Bin Li 等ICLR 2020 · 被引用 1,825 次
- Watch, Listen and Tell: Multi-Modal Weakly Supervised Dense Event CaptioningTanzila Rahman, Bicheng Xu, Leonid SigalICCV 2019 · 被引用 89 次
相关 Paper
- Damage Analysis via Bidirectional Multi-Task Cascaded Multimodal FusionTao Liang, Siying Wu, Junfeng Fang, Guowu Yang 等WWW 2025 · 被引用 1 次
- Multi-modal Graph Fusion for Named Entity Recognition with Targeted Visual GuidanceDong Zhang, Suzhong Wei, Shoushan Li, Hanqian Wu 等AAAI 2021 · 被引用 240 次
- CrisisTS: Coupling Social Media Textual Data and Meteorological Time Series for Urgency ClassificationRomain Meunier, Farah Benamara, Véronique Moriceau, Zhongzheng Qiao 等ACL 2025 · 被引用 1 次
- Enhancing Fake News Detection in Social Media via Label Propagation on Cross-modal Tweet GraphWanqing Zhao, Yuta Nakashima, Haiyuan Chen, Noboru BabaguchiACM MM 2023 · 被引用 10 次
- MultiEMO: An Attention-Based Correlation-Aware Multimodal Fusion Framework for Emotion Recognition in ConversationsTao Shi, Shao-Lun HuangACL 2023 · 被引用 76 次
