Towards Semantic Consistency: Dirichlet Energy Driven Robust Multi-Modal Entity Alignment
Yuanyi Wang, Haifeng Sun, Jiabo Wang, Jingyu Wang, Wei Tang, Qi Qi, Shaoling Sun, Jianxin Liao
Abstract
Multi-Modal Entity Alignment (MMEA) is a pivotal task in Multi-Modal Knowledge Graphs (MMKGs), seeking to identify identical entities by leveraging associated modal attributes. However, real-world MMKGs confront the challenges of semantic inconsistency arising from diverse and incomplete data sources. This inconsistency is predominantly caused by the absence of specific modal attributes, manifesting in two distinct forms: disparities in attribute counts or the absence of certain modalities. Current methods address these issues through attribute interpolation, but their reliance on predefined distributions introduces modality noise, compromising original semantic information. Furthermore, the absence of a generalizable theoretical principle hampers progress towards achieving semantic consistency. In this work, we propose a generalizable theoretical principle by examining semantic consistency from the perspective of Dirichlet energy. Our research reveals that, in the presence of semantic inconsistency, models tend to overfit to modality noise, leading to over-smoothing and performance oscillations or declines, particularly in scenarios with a high rate of missing modality. To overcome these challenges, we propose DESAlign, a robust method addressing the over-smoothing caused by semantic inconsistency and interpolating missing semantics using existing modalities. Specifically, we devise a training strategy for multi-modal knowledge graph learning based on our proposed principle. Then, we introduce a propagation strategy that utilizes existing features to provide interpolation solutions for missing semantic features. DESAlign outperforms existing approaches across 60 benchmark splits, encompassing both monolingual and bilingual scenarios, achieving state-of-the-art performance. Experiments on splits with high missing modal attributes demonstrate its effectiveness, providing a robust MMEA solution to semantic inconsistency in real-world MMKGs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 452dee72-4fb4-4aae-8efe-4b25aeee4844Cited by top-tier papers7
- InfiGFusion: Graph-on-Logits Distillation via Efficient Gromov-Wasserstein for Model FusionYuanyi Wang, Zhaoyi Yan, Yiming Zhang, Qi Zhou et al.NeurIPS 2025 · 12 citations
- Mitigating Modality Bias in Multi-modal Entity Alignment from a Causal PerspectiveTaoyu Su, Jiawei Sheng, Duohe Ma, Xiaodong Li et al.SIGIR 2025 · 4 citations
- PSQE: A Theoretical-Practical Approach to Pseudo Seed Quality Enhancement for Unsupervised Multimodal Entity AlignmentYunpeng Hong, Chenyang Bu, Jie Zhang, Yi He et al.KDD 2026
- MyGram: Modality-aware Graph Transformer with Global Distribution for Multi-modal Entity AlignmentZhifei Li, Ziyue Qin, Xiangyu Luo, Xiaoju Hou et al.AAAI 2026
- FSD-CAP: Fractional Subgraph Diffusion with Class-Aware Propagation for Graph Feature ImputationXin Qiao, Shijie Sun, Anqi Dong, Cong Hua et al.ICLR 2026
Builds on15
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Knowledge Graph Alignment Network with Gated Multi-Hop Neighborhood AggregationZequn Sun, Chengming Wang, Wei Hu, Muhao Chen et al.AAAI 2020 · 379 citations
- Dirichlet Energy Constrained Learning for Deep Graph Neural NetworksKaixiong Zhou, Xiao Huang, Daochen Zha, Rui Chen et al.NeurIPS 2021 · 171 citations
- Visual Pivoting for (Unsupervised) Entity AlignmentFangyu Liu, Muhao Chen, Dan Roth, Nigel CollierAAAI 2021 · 159 citations
- Exploring and Evaluating Attributes, Values, and Structures for Entity AlignmentZhiyuan Liu, Yixin Cao, Liangming Pan, Juanzi Li et al.EMNLP 2020 · 110 citations
Related papers
- Attribute-Consistent Knowledge Graph Representation Learning for Multi-Modal Entity AlignmentQian Li, Shu Guo, Yangyifei Luo, Cheng Ji et al.WWW 2023 · 56 citations
- Learning with Dual-level Noisy Correspondence for Multi-modal Entity AlignmentHaobin Li, Yijie Lin, Peng Hu, Mouxing Yang et al.ICLR 2026 · 2 citations
- Breaking the Noise Barrier: LLM-Guided Semantic Filtering and Enhancement for Multi-Modal Entity AlignmentChenglong Lu, Chenxiao Li, Jingwei Cheng, Yongquan Ji et al.EMNLP 2025
- Multi-modal Siamese Network for Entity AlignmentLiyi Chen, Zhi Li, Tong Xu, Han Wu et al.KDD 2022 · 82 citations
- IBMEA: Exploring Variational Information Bottleneck for Multi-modal Entity AlignmentTaoyu Su, Jiawei Sheng, Shicheng Wang, Xinghua Zhang et al.ACM MM 2024 · 7 citations
