Multimodal Contextual Interactions of Entities: A Modality Circular Fusion Approach for Link Prediction
Jing Yang, Shundong Yang, Yuan Gao, Jieming Yang, Laurence T. Yang
Abstract
Link prediction aims to infer missing valid triplets to complete knowledge graphs, with recent inclusion of multimodal information to enrich entity representations. Existing methods project multimodal information into a unified embedding space or learn modality-specific features separately for later integration. However, performance was limited in such studies due to neglecting the modalities compatibility and conflict semantic carried by entities in valid and invalid triplets. In this paper, we aim at modeling inter-entity modality interactions and thus propose a novel Modality Circular fusion approach (MoCi), which interweaves multimodal contextual of entities. Firstly, unlike most methods in this task that directly fuse modalities, we design a triplets-prompt modality contrastive pre-training to align modality semantics beforehand. Moreover, we propose a modality circular fusion model using a simple yet efficient multilinear transformation strategy. This allows explicit inter-entity modality interactions, distinguishing it from methods confined to fuse within individual entities. To the best of our knowledge, MoCi presents one of the pioneering frameworks that tailored to grasp inter-entity modality semantics for better link prediction. Extensive experiments on seven datasets demonstrate our model yields SOTA performance, confirming the efficacy of MoCi in modeling inter-entity modality interactions. Our code is released at https://github.com/MoCiGitHub/MoCi.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 270cd83b-9b3d-4e42-ad02-cc1ec1d6abc7Cited by top-tier papers1
Ask how each one uses itRelated papers
- IMF: Interactive Multimodal Fusion Model for Link PredictionXinhang Li, Xiangyu Zhao, Jiaxing Xu, Yong Zhang et al.WWW 2023 · 113 citations
- HyperFM: Fact-Centric Multimodal Fusion for Link Prediction over Hyper-Relational Knowledge GraphsYuhuan Lu, Weijian Yu, Xin Jing, Dingqi YangACL 2025 · 2 citations
- Mixed-Curvature Multi-Modal Knowledge Graph CompletionYuxiao Gao, Fuwei Zhang, Zhao Zhang, Xiaoshuang Min et al.AAAI 2025 · 5 citations
- LBMKGC: Large Model-Driven Balanced Multimodal Knowledge Graph CompletionYuan Guo, Qian Ma, Hui Li, Qiao Ning et al.NeurIPS 2025 · 3 citations
- Multi-modal Siamese Network for Entity AlignmentLiyi Chen, Zhi Li, Tong Xu, Han Wu et al.KDD 2022 · 82 citations
