Multimodal Dynamics: Dynamical Fusion for Trustworthy Multimodal Classification
Zongbo Han, Fan Yang, Junzhou Huang, Changqing Zhang, Jianhua Yao
摘要
Integration of heterogeneous and high-dimensional data (e.g., multiomics) is becoming increasingly important. Existing multimodal classification algorithms mainly focus on improving performance by exploiting the complementarity from different modalities. However, conventional approaches are basically weak in providing trustworthy multimodal fusion, especially for safety-critical applications (e.g., medical diagnosis). For this issue, we propose a novel trustworthy multimodal classification algorithm termed Multimodal Dynamics, which dynamically evaluates both the feature-level and modality-level informativeness for different samples and thus trustworthily integrates multiple modalities. Specifically, a sparse gating is introduced to capture the information variation of each within-modality feature and the true class probability is employed to assess the classification confidence of each modality. Then a transparent fusion algorithm based on the dynamical informativeness estimation strategy is induced. To the best of our knowledge, this is the first work to jointly model both feature and modality variation for different samples to provide trustworthy fusion in multi-modal classification. Extensive experiments are conducted on multimodal medical classification datasets. In these experiments, superior performance and trustworthiness of our algorithm are clearly validated compared to the state-of-the-art methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper25
- Provable Dynamic Fusion for Low-Quality Multimodal DataQingyang Zhang, Haitao Wu, Changqing Zhang, Qinghua Hu 等ICML 2023 · 被引用 143 次
- Revisiting Disentanglement and Fusion on Modality and Context in Conversational Multimodal Emotion RecognitionBobo Li, Hao Fei, Lizi Liao, Yu Zhao 等ACM MM 2023 · 被引用 76 次
- Coupled Mamba: Enhanced Multimodal Fusion with Coupled State Space ModelWenbing Li, Hang Zhou, Junqing Yu, Zikai Song 等NeurIPS 2024 · 被引用 71 次
- DecAlign: Hierarchical Cross-Modal Alignment for Decoupled Multimodal Representation LearningChengxuan Qian, Shuo Xing, Li Li, Yue Zhao 等ICLR 2026 · 被引用 42 次
- Multi-Level Confidence Learning for Trustworthy Multimodal ClassificationXiao Zheng, Chang Tang, Zhiguo Wan, Chengyu Hu 等AAAI 2023 · 被引用 41 次
它引用的顶会 Paper13
- Uncertainty Estimation Using a Single Deep Deterministic Neural NetworkJoost van Amersfoort, Lewis Smith, Yee Whye Teh, Yarin GalICML 2020 · 被引用 529 次
- What Makes Multi-Modal Learning Better than Single (Provably)Yu Huang, Chenzhuang Du, Zihui Xue, Xuanyao Chen 等NeurIPS 2021 · 被引用 404 次
- UniT: Multimodal Multitask Learning with a Unified TransformerRonghang Hu, Amanpreet SinghICCV 2021 · 被引用 354 次
- Deep Multimodal Fusion by Channel ExchangingYikai Wang, Wenbing Huang, Fuchun Sun, Tingyang Xu 等NeurIPS 2020 · 被引用 321 次
- Training independent subnetworks for robust predictionMarton Havasi, Rodolphe Jenatton, Stanislav Fort, Jeremiah Zhe Liu 等ICLR 2021 · 被引用 235 次
相关 Paper
- DPNET: Dynamic Poly-attention Network for Trustworthy Multi-modal ClassificationXin Zou, Chang Tang, Xiao Zheng, Zhenglai Li 等ACM MM 2023 · 被引用 16 次
- Scalable Medical Multimodal Fusion via Symmetric Consistency ModelingXiaowen Sun, Hui Liu, Gongguan Chen, Ning MaoICML 2026
- Trusted Multi-View ClassificationZongbo Han, Changqing Zhang, Huazhu Fu, Joey Tianyi ZhouICLR 2021
- Trustworthy Multimodal Regression with Mixture of Normal-inverse Gamma DistributionsHuan Ma, Zongbo Han, Changqing Zhang, Huazhu Fu 等NeurIPS 2021 · 被引用 78 次
- Sparse Maximum Margin Learning from Multimodal Human Behavioral PatternsErvine Zheng, Qi Yu, Zhi ZhengAAAI 2023
