Breaking the Noise Barrier: LLM-Guided Semantic Filtering and Enhancement for Multi-Modal Entity Alignment
Chenglong Lu, Chenxiao Li, Jingwei Cheng, Yongquan Ji, Guoqing Chen, Fu Zhang
摘要
Multi-modal entity alignment (MMEA) aims to identify equivalent entities between two multimodal knowledge graphs (MMKGs). Existing methods have made substantial advancements in enhancing multi-modal fusion. However, the intrinsic noise within modalities, such as the inconsistency in visual modality and redundant attributes, has not been thoroughly investigated. Excessive noise not only weakens semantic representation but also increases the risk of overfitting in attention-based fusion methods. To address this, we propose LGEA (LLM-Guided Entity Alignment), a novel LLM-guided MMEA framework that prioritizes noise reduction before fusion. Specifically, LGEA introduces two key strategies: (1) fine-grained visual filtering to remove irrelevant images at the semantic level, and (2) contextual summarization of attribute information to enhance entity semantics. To our knowledge, we are the first work to apply LLMs for both visual filtering and attribute-level semantic enhancement in MMEA. Experiments on multiple benchmarks, including the noisy FBYG dataset, show that LGEA sets a new state-of-the-art (SOTA) in robust multi-modal alignment, highlighting the potential of noiseaware strategies as a promising direction for future MMEA research 1 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper5
- BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and GenerationJunnan Li, Dongxu Li, Caiming Xiong, Steven C. H. HoiICML 2022 · 被引用 6,549 次
- Visual Pivoting for (Unsupervised) Entity AlignmentFangyu Liu, Muhao Chen, Dan Roth, Nigel CollierAAAI 2021 · 被引用 159 次
- AFDGCF: Adaptive Feature De-correlation Graph Collaborative Filtering for RecommendationsWei Wu, Chao Wang, Dazhong Shen, Chuan Qin 等SIGIR 2024 · 被引用 33 次
- Tackling Uncertain Correspondences for Multi-Modal Entity AlignmentLiyi Chen, Ying Sun, Shengzhe Zhang, Yuyang Ye 等NeurIPS 2024 · 被引用 20 次
- SimDiff: Simple Denoising Probabilistic Latent Diffusion Model for Data Augmentation on Multi-modal Knowledge GraphRan Li, Shimin Di, Lei Chen, Xiaofang ZhouKDD 2024 · 被引用 5 次
相关 Paper
- Attribute-Consistent Knowledge Graph Representation Learning for Multi-Modal Entity AlignmentQian Li, Shu Guo, Yangyifei Luo, Cheng Ji 等WWW 2023 · 被引用 56 次
- Learning with Dual-level Noisy Correspondence for Multi-modal Entity AlignmentHaobin Li, Yijie Lin, Peng Hu, Mouxing Yang 等ICLR 2026 · 被引用 2 次
- Towards Semantic Consistency: Dirichlet Energy Driven Robust Multi-Modal Entity AlignmentYuanyi Wang, Haifeng Sun, Jiabo Wang, Jingyu Wang 等ICDE 2024 · 被引用 13 次
- IBMEA: Exploring Variational Information Bottleneck for Multi-modal Entity AlignmentTaoyu Su, Jiawei Sheng, Shicheng Wang, Xinghua Zhang 等ACM MM 2024 · 被引用 7 次
- Pseudo-Label Calibration Semi-supervised Multi-Modal Entity AlignmentLuyao Wang, Pengnian Qi, Xigang Bao, Chunlai Zhou 等AAAI 2024 · 被引用 21 次
