Coarse-to-Fine Highlighting: Reducing Knowledge Hallucination in Large Language Models
Qitan Lv, Jie Wang, Hanzhu Chen, Bin Li, Yongdong Zhang, Feng Wu
摘要
Generation of plausible but incorrect factual information, often termed hallucination, has attracted significant research interest. Retrieval-augmented language model (RALM) -- which enhances models with up-to-date knowledge -- emerges as a promising method to reduce hallucination. However, existing RALMs may instead exacerbate hallucination when retrieving lengthy contexts. To address this challenge, we propose COFT, a novel COarse-to-Fine highlighTing method to focus on different granularity-level key texts, thereby avoiding getting lost in lengthy contexts. Specifically, COFT consists of three components: recaller, scorer, and selector. First, recaller applies a knowledge graph to extract potential key entities in a given context. Second, scorer measures the importance of each entity by calculating its contextual weight. Finally, selector selects high contextual weight entities with a dynamic threshold algorithm and highlights the corresponding paragraphs, sentences, or words in a coarse-to-fine manner. Extensive experiments on the knowledge hallucination benchmark demonstrate the effectiveness of COFT, leading to a superior performance over in the F1 score metric. Moreover, COFT also exhibits remarkable versatility across various long-form tasks, such as reading comprehension and question answering.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- LogicTree: Improving Complex Reasoning of LLMs via Instantiated Multi-step Synthetic Logical DataZehao Wang, Lin F. Yang, Jie Wang, Kehan Wang 等NeurIPS 2025 · 被引用 5 次
- Efficient Heuristics Generation for Solving Combinatorial Optimization Problems Using Large Language ModelsXuan Wu, Di Wang, Chunguo Wu, Lijie Wen 等KDD 2025 · 被引用 5 次
- Beyond RAG vs. Long-Context: Learning Distraction-Aware Retrieval for Efficient Knowledge GroundingSeong-Woong Shim, Myunsoo Kim, Jae Hyeon Cho, Byung-Jun LeeICLR 2026 · 被引用 1 次
- Efficient OpAmp Adaptation for Zoom Attention to Golden ContextsHaoyuan Wu, Rui Ming, Haisheng Zheng, Zhuolun He 等ACL 2025 · 被引用 1 次
- Graph-constrained Reasoning: Faithful Reasoning on Knowledge Graphs with Large Language ModelsLinhao Luo, Zicheng Zhao, Gholamreza Haffari, Yuan-Fang Li 等ICML 2025
它引用的顶会 Paper34
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni 等NeurIPS 2020 · 被引用 19,162 次
- Self-Refine: Iterative Refinement with Self-FeedbackAman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan 等NeurIPS 2023 · 被引用 4,972 次
- Big Bird: Transformers for Longer SequencesManzil Zaheer, Guru Guruganesh, Kumar Avinava Dubey, Joshua Ainslie 等NeurIPS 2020 · 被引用 3,159 次
相关 Paper
- Post-hoc Utterance Refining Method by Entity Mining for Faithful Knowledge Grounded ConversationsYoonna Jang, Suhyune Son, Jeongwoo Lee, Junyoung Son 等EMNLP 2023
- Micro-Macro Retrieval: Reducing Long-Form Hallucination in Large Language ModelsYujie Feng, Jian Li, Zhihan Zhou, Pengfei Xu 等ICLR 2026
- Boosting Retrieval-Augmented Generation with Generation-Augmented Retrieval: A Co-Training ApproachYubao Tang, Ruqing Zhang, Jiafeng Guo, Maarten de Rijke 等SIGIR 2025 · 被引用 2 次
- Mitigating Large Language Model Hallucinations via Autonomous Knowledge Graph-Based RetrofittingXinyan Guan, Yanjiang Liu, Hongyu Lin, Yaojie Lu 等AAAI 2024 · 被引用 127 次
- Chain-of-Note: Enhancing Robustness in Retrieval-Augmented Language ModelsWenhao Yu, Hongming Zhang, Xiaoman Pan, Peixin Cao 等EMNLP 2024 · 被引用 30 次
