Towards Faithfulness in Open Domain Table-to-text Generation from an Entity-centric View
Tianyu Liu, Xin Zheng, Baobao Chang, Zhifang Sui
摘要
In open domain table-to-text generation, we notice that the unfaithful generation usually contains hallucinated content which can not be aligned to any input table record. We thus try to evaluate the generation faithfulness with two entity-centric metrics: table record coverage and the ratio of hallucinated entities in text, both of which are shown to have strong agreement with human judgements. Then based on these metrics, we quantitatively analyze the correlation between training data quality and generation fidelity which indicates the potential usage of entity information in faithful generation. Motivated by these findings, we propose two methods for faithful generation: 1) augmented training by incorporating the auxiliary entity information, including both an augmented planbased model and an unsupervised model and 2) training instance selection based on faithfulness ranking. We show these approaches improve generation fidelity in both full dataset setting and few shot learning settings by both automatic and human evaluations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Factuality Enhanced Language Models for Open-Ended Text GenerationNayeon Lee, Wei Ping, Peng Xu, Mostofa Patwary 等NeurIPS 2022 · 被引用 318 次
- A Token-level Reference-free Hallucination Detection Benchmark for Free-form Text GenerationTianyu Liu, Yizhe Zhang, Chris Brockett, Yi Mao 等ACL 2022 · 被引用 194 次
- Mitigating Large Language Model Hallucinations via Autonomous Knowledge Graph-Based RetrofittingXinyan Guan, Yanjiang Liu, Hongyu Lin, Yaojie Lu 等AAAI 2024 · 被引用 127 次
- AutoDDG: Automated Dataset Description Generation using Large Language ModelsHaoxiang Zhang, Yurong Liu, Aécio S. R. Santos, Wei-Lun Hung 等SIGMOD 2026 · 被引用 17 次
- Truth-Conditional Captions for Time Series DataHarsh Jhamtani, Taylor Berg-KirkpatrickEMNLP 2021
它引用的顶会 Paper6
- Asking and Answering Questions to Evaluate the Factual Consistency of SummariesAlex Wang, Kyunghyun Cho, Mike LewisACL 2020 · 被引用 317 次
- ToTTo: A Controlled Table-To-Text Generation DatasetAnkur P. Parikh, Xuezhi Wang, Sebastian Gehrmann, Manaal Faruqui 等EMNLP 2020 · 被引用 69 次
- On Faithfulness and Factuality in Abstractive SummarizationJoshua Maynez, Shashi Narayan, Bernd Bohnet, Ryan T. McDonaldACL 2020 · 被引用 54 次
- Neural Data-to-Text Generation via Jointly Learning the Segmentation and CorrespondenceXiaoyu Shen, Ernie Chang, Hui Su, Cheng Niu 等ACL 2020 · 被引用 46 次
- Variational Template Machine for Data-to-Text GenerationRong Ye, Wenxian Shi, Hao Zhou, Zhongyu Wei 等ICLR 2020 · 被引用 45 次
相关 Paper
- R2D2: Robust Data-to-Text with Replacement DetectionLinyong Nan, Lorenzo Jaime Yu Flores, Yilun Zhao, Yixin Liu 等EMNLP 2022 · 被引用 10 次
- Towards Faithful Neural Table-to-Text Generation with Content-Matching ConstraintsZhenyi Wang, Xiaoyang Wang, Bang An, Dong Yu 等ACL 2020 · 被引用 21 次
- Post-hoc Utterance Refining Method by Entity Mining for Faithful Knowledge Grounded ConversationsYoonna Jang, Suhyune Son, Jeongwoo Lee, Junyoung Son 等EMNLP 2023
- SCOPE: A Self-supervised Framework for Improving Faithfulness in Conditional Text GenerationSong Duong, Florian Le Bronnec, Alexandre Allauzen, Vincent Guigue 等ICLR 2025
- Logical Natural Language Generation from Open-Domain TablesWenhu Chen, Jianshu Chen, Yu Su, Zhiyu Chen 等ACL 2020 · 被引用 116 次
