Entity Alignment with Noisy Annotations from Large Language Models
Shengyuan Chen, Qinggang Zhang, Junnan Dong, Wen Hua, Qing Li, Xiao Huang
Abstract
Entity alignment (EA) aims to merge two knowledge graphs (KGs) by identifying equivalent entity pairs. While existing methods heavily rely on human-generated labels, it is prohibitively expensive to incorporate cross-domain experts for annotation in real-world scenarios. The advent of Large Language Models (LLMs) presents new avenues for automating EA with annotations, inspired by their comprehensive capability to process semantic information. However, it is nontrivial to directly apply LLMs for EA since the annotation space in real-world KGs is large. LLMs could also generate noisy labels that may mislead the alignment. To this end, we propose a unified framework, LLM4EA, to effectively leverage LLMs for EA. Specifically, we design a novel active learning policy to significantly reduce the annotation space by prioritizing the most valuable entities based on the entire inter-KG and intra-KG structure. Moreover, we introduce an unsupervised label refiner to continuously enhance label accuracy through in-depth probabilistic reasoning. We iteratively optimize the policy based on the feedback from a base EA model. Extensive experiments demonstrate the advantages of LLM4EA on four benchmark datasets in terms of effectiveness, robustness, and efficiency. Codes are available via https://github.com/chensyCN/llm4ea_official.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers26
- When to use Graphs in RAG: A Comprehensive Analysis for Graph Retrieval-Augmented GenerationZhishang Xiang, Chuanjie Wu, Qinggang Zhang, Shengyuan Chen et al.ICLR 2026 · 56 citations
- LinearRAG: Linear Graph Retrieval Augmented Generation on Large-scale CorporaLuyao Zhuang, Shengyuan Chen, Yilin Xiao, Huachi Zhou et al.ICLR 2026 · 54 citations
- ZeroG: Investigating Cross-dataset Zero-shot Transferability in GraphsYuhan Li, Peisong Wang, Zhixun Li, Jeffrey Xu Yu et al.KDD 2024 · 19 citations
- FaithfulRAG: Fact-Level Conflict Modeling for Context-Faithful Retrieval-Augmented GenerationQinggang Zhang, Zhishang Xiang, Yilin Xiao, Le Wang et al.ACL 2025 · 18 citations
- You Don't Need Pre-Built Graphs for RAG: Retrieval Augmented Generation with Adaptive Reasoning StructuresShengyuan Chen, Chuang Zhou, Zheng Yuan, Qinggang Zhang et al.AAAI 2026 · 14 citations
Builds on13
- A Benchmarking Study of Embedding-based Entity Alignment for Knowledge GraphsZequn Sun, Qingheng Zhang, Wei Hu, Chengming Wang et al.VLDB 2020 · 297 citations
- Knowledge Graph Prompting for Multi-Document Question AnsweringYu Wang, Nedim Lipka, Ryan A. Rossi, Alexa F. Siu et al.AAAI 2024 · 290 citations
- Boosting the Speed of Entity Alignment 10 ×: Dual Attention Matching Network with Normalized Hard Sample MiningXin Mao, Wenting Wang, Yuanbin Wu, Man LanWWW 2021 · 148 citations
- Efficient Probabilistic Logic Reasoning with Graph Neural NetworksYuyu Zhang, Xinshi Chen, Yuan Yang, Arun Ramamurthy et al.ICLR 2020 · 119 citations
- Label-free Node Classification on Graphs with Large Language Models (LLMs)Zhikai Chen, Haitao Mao, Hongzhi Wen, Haoyu Han et al.ICLR 2024 · 103 citations
Related papers
- HLMEA: Unsupervised Entity Alignment Based on Hybrid Language ModelsXiongnan Jin, Zhilin Wang, Jinpeng Chen, Liu Yang et al.AAAI 2025 · 5 citations
- EA-Agent: A Structured Multi-Step Reasoning Agent for Entity AlignmentYixuan Nan, Xixun Lin, Yanmin Shang, Ge Zhang et al.ACL 2026
- Multi-Modal Fact Knowledge Generation for Imbalanced Cross-Source Entity AlignmentQian Li, Cheng Ji, Zhaoji Liang, Yuzheng Zhang et al.AAAI 2026
- ZeroEA: A Zero-Training Entity Alignment Framework via Pre-Trained Language ModelNan Huo, Reynold Cheng, Ben Kao, Wentao Ning et al.VLDB 2024 · 16 citations
- ActiveEA: Active Learning for Neural Entity AlignmentBing Liu, Harrisen Scells, Guido Zuccon, Wen Hua et al.EMNLP 2021
