HGMF: Heterogeneous Graph-based Fusion for Multimodal Data with Incompleteness
Jiayi Chen, Aidong Zhang
Abstract
With the advances in data collection techniques, large amounts of multimodal data collected from multiple sources are becoming available. Such multimodal data can provide complementary information that can reveal fundamental characteristics of real-world subjects. Thus, multimodal machine learning has become an active research area. Extensive works have been developed to exploit multimodal interactions and integrate multi-source information. However, multimodal data in the real world usually comes with missing modalities due to various reasons, such as sensor damage, data corruption, and human mistakes in recording. Effectively integrating and analyzing multimodal data with incompleteness remains a challenging problem. We propose a Heterogeneous Graph-based Multimodal Fusion (HGMF) approach to enable multimodal fusion of incomplete data within a heterogeneous graph structure. The proposed approach develops a unique strategy for learning on incomplete multimodal data without data deletion or data imputation. More specifically, we construct a heterogeneous hypernode graph to model the multimodal data having different combinations of missing modalities, and then we formulate a graph neural network based transductive learning framework to project the heterogeneous incomplete data onto a unified embedding space, and multi-modalities are fused along the way. The learning framework captures modality interactions from available data, and leverages the relationships between different incompleteness patterns. Our experimental results demonstrate that the proposed method outperforms existing graph-based as well as non-graph based baselines on three different datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers16
- M3Care: Learning with Missing Modalities in Multimodal Healthcare DataChaohe Zhang, Xu Chu, Liantao Ma, Yinghao Zhu et al.KDD 2022 · 78 citations
- Multimodal Patient Representation Learning with Missing Modalities and LabelsZhenbang Wu, Anant Dadu, Nicholas J. Tustison, Brian B. Avants et al.ICLR 2024 · 38 citations
- Knowledge Perceived Multi-modal Pretraining in E-commerceYushan Zhu, Huaixiao Zhao, Wen Zhang, Ganqiang Ye et al.ACM MM 2021 · 21 citations
- On Disentanglement of Asymmetrical Knowledge Transfer for Modality-Task Agnostic Federated LearningJiayi Chen, Aidong ZhangAAAI 2024 · 18 citations
- Balancing Similarity and Complementarity for Federated LearningKunda Yan, Sen Cui, Abudukelimu Wuerkaixi, Jingfeng Zhang et al.ICML 2024 · 13 citations
Related papers
- Label-Semantics-Guided Multi-View Multi-Label Learning via High-Order Semantic FusionKaixiang Wang, Xiaojian Ding, Wanqi Yang, Ming YangACM MM 2025
- Modality to Modality Translation: An Adversarial Representation Learning and Graph Fusion Network for Multimodal FusionSijie Mai, Haifeng Hu, Songlong XingAAAI 2020 · 233 citations
- Heterogeneous Graph Neural Network via Attribute CompletionDi Jin, Cuiying Huo, Chundong Liang, Liang YangWWW 2021 · 220 citations
- Multimodal Learning with Incomplete Modalities by Knowledge DistillationQi Wang, Liang Zhan, Paul M. Thompson, Jiayu ZhouKDD 2020 · 80 citations
- Hyper-Modality Enhancement for Multimodal Sentiment Analysis with Missing ModalitiesYan Zhuang, Minhao Liu, Wei Bai, Yanru Zhang et al.NeurIPS 2025 · 10 citations
