IMF: Interactive Multimodal Fusion Model for Link Prediction
Xinhang Li, Xiangyu Zhao, Jiaxing Xu, Yong Zhang, Chunxiao Xing
Abstract
Link prediction aims to identify potential missing triples in knowledge graphs. To get better results, some recent studies have introduced multimodal information to link prediction. However, these methods utilize multimodal information separately and neglect the complicated interaction between different modalities. In this paper, we aim at better modeling the inter-modality information and thus introduce a novel Interactive Multimodal Fusion (IMF) model to integrate knowledge from different modalities. To this end, we propose a two-stage multimodal fusion framework to preserve modality-specific knowledge as well as take advantage of the complementarity between different modalities. Instead of directly projecting different modalities into a unified space, our multimodal fusion module limits the representations of different modalities independent while leverages bilinear pooling for fusion and incorporates contrastive learning as additional constraints. Furthermore, the decision fusion module delivers the learned weighted average over the predictions of all modalities to better incorporate the complementarity of different modalities. Our approach has been demonstrated to be effective through empirical evaluations on several real-world datasets. The implementation code is available online at https://github.com/HestiaSky/IMF-Pytorch .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 236884db-c789-490e-9a54-ccc08b5ce7d2Cited by top-tier papers15
- NativE: Multi-modal Knowledge Graph Completion in the WildYichi Zhang, Zhuo Chen, Lingbing Guo, Yajing Xu et al.SIGIR 2024 · 39 citations
- Tokenization, Fusion, and Augmentation: Towards Fine-grained Multi-modal Entity RepresentationYichi Zhang, Zhuo Chen, Lingbing Guo, Yajing Xu et al.AAAI 2025 · 25 citations
- APKGC: Noise-enhanced Multi-Modal Knowledge Graph Completion with Attention PenaltyYue Jian, Xiangyu Luo, Zhifei Li, Miao Zhang et al.AAAI 2025 · 21 citations
- Contrast then Memorize: Semantic Neighbor Retrieval-Enhanced Inductive Multimodal Knowledge Graph CompletionYu Zhao, Ying Zhang, Baohang Zhou, Xinying Qian et al.SIGIR 2024 · 15 citations
- Mixed-Curvature Multi-Modal Knowledge Graph CompletionYuxiao Gao, Fuwei Zhang, Zhao Zhang, Xiaoshuang Min et al.AAAI 2025 · 5 citations
Builds on10
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Align before Fuse: Vision and Language Representation Learning with Momentum DistillationJunnan Li, Ramprasaath R. Selvaraju, Akhilesh Gotmare, Shafiq R. Joty et al.NeurIPS 2021 · 2,985 citations
- Learning Hierarchy-Aware Knowledge Graph Embeddings for Link PredictionZhanqiu Zhang, Jianyu Cai, Yongdong Zhang, Jie WangAAAI 2020 · 481 citations
- InteractE: Improving Convolution-Based Knowledge Graph Embeddings by Increasing Feature InteractionsShikhar Vashishth, Soumya Sanyal, Vikram Nitin, Nilesh Agrawal et al.AAAI 2020 · 393 citations
- AutoSF: Searching Scoring Functions for Knowledge Graph EmbeddingYongqi Zhang, Quanming Yao, Wenyuan Dai, Lei ChenICDE 2020 · 89 citations
Related papers
- Multimodal Contextual Interactions of Entities: A Modality Circular Fusion Approach for Link PredictionJing Yang, Shundong Yang, Yuan Gao, Jieming Yang et al.ACM MM 2024 · 7 citations
- WFF: Wavelet-based Information Fusion for Multimodal Knowledge Graph Link PredictionXiaodi Xu, Lijie Li, Ye Wang, Tao Ren et al.ACM MM 2025 · 1 citation
- MoSE: Modality Split and Ensemble for Multimodal Knowledge Graph CompletionYu Zhao, Xiangrui Cai, Yike Wu, Haiwei Zhang et al.EMNLP 2022 · 66 citations
- Multimodal Entity Linking with Gated Hierarchical Fusion and Contrastive TrainingPeng Wang, Jiangheng Wu, Xiaohang ChenSIGIR 2022 · 52 citations
- Enhancing Multi-Modal Entity Alignment via Multi-Grained Decision FusionYu Xing, Qizhuo Xie, You Lv, Ziyang Zhou et al.WWW 2026
