Hallucination Mitigation in Natural Language Generation from Large-Scale Open-Domain Knowledge Graphs
Xiao Shi, Zhengyuan Zhu, Zeyu Zhang, Chengkai Li
摘要
In generating natural language descriptions for knowledge graph triples, prior works used either small-scale, human-annotated datasets or datasets with limited variety of graph shapes, e.g., those having mostly star graphs. Graph-to-text models trained and evaluated on such datasets are largely not assessed for more realistic large-scale, open-domain settings. We introduce a new dataset, GraphNarrative, to fill this gap. Fine-tuning transformer-based pre-trained language models has achieved state-of-the-art performance among graph-to-text models. However, this method suffers from information hallucination—the generated text may contain fabricated facts not present in input graphs. We propose a novel approach that, given a graph-sentence pair in GraphNarrative, trims the sentence to eliminate portions that are not present in the corresponding graph, by utilizing the sentence’s dependency parse tree. Our experiment results verify this approach using models trained on GraphNarrative and existing datasets. The dataset, source code, and trained models are released at https://github.com/idirlab/graphnarrator.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Debate on Graph: A Flexible and Reliable Reasoning Framework for Large Language ModelsJie Ma, Zhitao Gao, Qi Chai, Wangchun Sun 等AAAI 2025 · 被引用 8 次
- An Audit on the Perspectives and Challenges of Hallucinations in NLPPranav Narayanan Venkit, Tatiana Chakravorti, Vipul Gupta, Heidi Biggs 等EMNLP 2024 · 被引用 8 次
它引用的顶会 Paper4
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- HTLM: Hyper-Text Pre-Training and Prompting of Language ModelsArmen Aghajanyan, Dmytro Okhonko, Mike Lewis, Mandar Joshi 等ICLR 2022 · 被引用 82 次
- Open Domain Question Answering with A Unified Knowledge InterfaceKaixin Ma, Hao Cheng, Xiaodong Liu, Eric Nyberg 等ACL 2022 · 被引用 45 次
相关 Paper
- Structure-aware Knowledge Graph-to-text Generation with Planning Selection and Similarity DistinctionFeng Zhao, Hongzhi Zou, Cheng YanEMNLP 2023 · 被引用 5 次
- ScriptWriter: Narrative-Guided Script GenerationYutao Zhu, Ruihua Song, Zhicheng Dou, Jian-Yun Nie 等ACL 2020 · 被引用 23 次
- DiscoSG: Towards Discourse-Level Text Scene Graph Parsing through Iterative Graph RefinementShaoqing Lin, Chong Teng, Fei Li, Donghong Ji 等EMNLP 2025
- Generating Coherent Narratives by Learning Dynamic and Discrete Entity States with a Contrastive FrameworkJian Guan, Zhenyu Yang, Rongsheng Zhang, Zhipeng Hu 等AAAI 2023 · 被引用 11 次
- ENT-DESC: Entity Description Generation by Exploring Knowledge GraphLiying Cheng, Dekun Wu, Lidong Bing, Yan Zhang 等EMNLP 2020 · 被引用 22 次
