Cross-Modal Graph Attention Network for Entity Alignment
Baogui Xu, Chengjin Xu, Bing Su
Abstract
The increasing popularity of multi-modal knowledge graphs (MMKGs) has led to a need for efficient entity alignment techniques that can exploit multi-modal information to integrate knowledge from different sources. GNN-based multi-modal entity alignment (MMEA) methods have achieved significant progress in entity alignment(EA) areas. However, these methods only rely on Graph Neural Networks (GNNs) to encode structural information, while ignoring visual and semantic modalities, which may lead to incomplete representation, thus how to integrate the visual and semantic information into GNN-based EA methods remains unexplored. In light of our insight that incorporating the message-passing mechanism of Graph Neural Networks to integrate multi-modal information is essential for fully exploiting the graph representation capability of GNN, we propose a novel Cross-modal Graph attention network for Entity Alignment (XGEA) that enables visual knowledge to interact with other views of the entity, including structural and literal information. We leverage the information from one modality as complementary relation information to compute the attention of another modality in the graph attention layers, enabling the learning of entity embedding by integrating multiple modalities. Moreover, the quantity of labeled data plays a crucial role in model performance, yet obtaining sufficient training data is expensive. To mitigate this issue, we use visual and semantic information to generate pseudo-pairs and propose a soft pseudo-labeling method for entity alignment to assign weights to the augmented training data to balance its quantity and quality. Extensive experiments show that our XGEA achieves superior performance consistently over the state-of-the-art MMEA baselines.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 06dbb91c-00e7-4ea0-af5a-3b2ed2cce023Cited by top-tier papers5
- IBMEA: Exploring Variational Information Bottleneck for Multi-modal Entity AlignmentTaoyu Su, Jiawei Sheng, Shicheng Wang, Xinghua Zhang et al.ACM MM 2024 · 7 citations
- HLMEA: Unsupervised Entity Alignment Based on Hybrid Language ModelsXiongnan Jin, Zhilin Wang, Jinpeng Chen, Liu Yang et al.AAAI 2025 · 5 citations
- Mitigating Modality Bias in Multi-modal Entity Alignment from a Causal PerspectiveTaoyu Su, Jiawei Sheng, Duohe Ma, Xiaodong Li et al.SIGIR 2025 · 4 citations
- Learning with Dual-level Noisy Correspondence for Multi-modal Entity AlignmentHaobin Li, Yijie Lin, Peng Hu, Mouxing Yang et al.ICLR 2026 · 2 citations
- Implicit Fine-tuning via Context Engineering: A Curriculum Learning Framework for Multimodal Entity AlignmentYunpeng Hong, Chenyang Bu, Di Wu, Yi He et al.KDD 2026
Related papers
- Multi-modal Siamese Network for Entity AlignmentLiyi Chen, Zhi Li, Tong Xu, Han Wu et al.KDD 2022 · 82 citations
- Tackling Uncertain Correspondences for Multi-Modal Entity AlignmentLiyi Chen, Ying Sun, Shengzhe Zhang, Yuyang Ye et al.NeurIPS 2024 · 20 citations
- Explicit-Implicit Entity Alignment Method in Multi-modal Knowledge GraphsLuyao Wang, Chunlai Zhou, Biao QinKDD 2025
- Pseudo-Label Calibration Semi-supervised Multi-Modal Entity AlignmentLuyao Wang, Pengnian Qi, Xigang Bao, Chunlai Zhou et al.AAAI 2024 · 21 citations
- Attribute-Consistent Knowledge Graph Representation Learning for Multi-Modal Entity AlignmentQian Li, Shu Guo, Yangyifei Luo, Cheng Ji et al.WWW 2023 · 56 citations
