Multimodal Categorization of Crisis Events in Social Media
Mahdi Abavisani, Liwei Wu, Shengli Hu, Joel R. Tetreault, Alejandro Jaimes
Abstract
Recent developments in image classification and natural language processing, coupled with the rapid growth in social media usage, have enabled fundamental advances in detecting breaking events around the world in real-time. Emergency response is one such area that stands to gain from these advances. By processing billions of texts and images a minute, events can be automatically detected to enable emergency response workers to better assess rapidly evolving situations and deploy resources accordingly. To date, most event detection techniques in this area have focused on image-only or text-only approaches, limiting detection performance and impacting the quality of information delivered to crisis response teams. In this paper, we present a new multimodal fusion method that leverages both images and texts as input. In particular, we introduce a cross-attention module that can filter uninformative and misleading components from weak modalities on a sample by sample basis. In addition, we employ a multimodal graph-based approach to stochastically transition between embeddings of different multimodal pairs during training to better regularize the learning process as well as dealing with limited training data by constructing new matched pairs from different samples. We show that our method outperforms the unimodal approaches and strong multimodal baselines by a large margin on three crisis-related tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dcfc3c78-248b-404c-bf2a-1939bb86dba9Cited by top-tier papers12
- Graph Neural Architecture Search Under Distribution ShiftsYijian Qin, Xin Wang, Ziwei Zhang, Pengtao Xie et al.ICML 2022 · 41 citations
- Expanding Large Pre-trained Unimodal Models with Multimodal Information Injection for Image-Text Multimodal ClassificationTao Liang, Guosheng Lin, Mingyang Wan, Tianrui Li et al.CVPR 2022 · 39 citations
- Aligning Vision to Language: Annotation-Free Multimodal Knowledge Graph Construction for Enhanced LLMs ReasoningJunming Liu, Siyuan Meng, Yanting Gao, Song Mao et al.ICCV 2025 · 34 citations
- Multimodal Graph Neural Architecture Search under Distribution ShiftsJie Cai, Xin Wang, Haoyang Li, Ziwei Zhang et al.AAAI 2024 · 20 citations
- Natural Disaster Tweets Classification Using Multimodal DataMohammad Basit, Bashir Alam, Zubaida Fatima, Salman ShaikhEMNLP 2023 · 12 citations
Builds on3
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel et al.ICLR 2020 · 7,418 citations
- VL-BERT: Pre-training of Generic Visual-Linguistic RepresentationsWeijie Su, Xizhou Zhu, Yue Cao, Bin Li et al.ICLR 2020 · 1,825 citations
- Watch, Listen and Tell: Multi-Modal Weakly Supervised Dense Event CaptioningTanzila Rahman, Bicheng Xu, Leonid SigalICCV 2019 · 89 citations
Related papers
- Damage Analysis via Bidirectional Multi-Task Cascaded Multimodal FusionTao Liang, Siying Wu, Junfeng Fang, Guowu Yang et al.WWW 2025 · 1 citation
- Multi-modal Graph Fusion for Named Entity Recognition with Targeted Visual GuidanceDong Zhang, Suzhong Wei, Shoushan Li, Hanqian Wu et al.AAAI 2021 · 240 citations
- CrisisTS: Coupling Social Media Textual Data and Meteorological Time Series for Urgency ClassificationRomain Meunier, Farah Benamara, Véronique Moriceau, Zhongzheng Qiao et al.ACL 2025 · 1 citation
- Enhancing Fake News Detection in Social Media via Label Propagation on Cross-modal Tweet GraphWanqing Zhao, Yuta Nakashima, Haiyuan Chen, Noboru BabaguchiACM MM 2023 · 10 citations
- MultiEMO: An Attention-Based Correlation-Aware Multimodal Fusion Framework for Emotion Recognition in ConversationsTao Shi, Shao-Lun HuangACL 2023 · 76 citations
