Dialogues Are Not Just Text: Modeling Cognition for Dialogue Coherence Evaluation
Xue Li, Jia Su, Yang Yang, Zipeng Gao, Xinyu Duan, Yi Guan
摘要
The generation of logically coherent dialogues by humans relies on underlying cognitive abilities. Based on this, we redefine the dialogue coherence evaluation process, combining cognitive judgment with the basic text to achieve a more human-like evaluation. We propose a novel dialogue evaluation framework based on Dialogue Cognition Graph (DCGEval) to implement the fusion by in-depth interaction between cognition modeling and text modeling. The proposed Abstract Meaning Representation (AMR) based graph structure called DCG aims to uniformly model four dialogue cognitive abilities. Specifically, core-semantic cognition is modeled by converting the utterance into an AMR graph, which can extract essential semantic information without redundancy. The temporal and role cognition are modeled by establishing logical relationships among the different AMR graphs. Finally, the commonsense knowledge from Concept-Net is fused to express commonsense cognition. Experiments demonstrate the necessity of modeling human cognition for dialogue evaluation, and our DCGEval presents stronger correlations with human judgments compared to other state-ofthe-art evaluation metrics.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper10
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger 等ICLR 2020 · 被引用 8,443 次
- GreaseLM: Graph REASoning Enhanced Language ModelsXikun Zhang, Antoine Bosselut, Michihiro Yasunaga, Hongyu Ren 等ICLR 2022 · 被引用 285 次
- AMR Parsing via Graph-Sequence Iterative InferenceDeng Cai, Wai LamACL 2020 · 被引用 83 次
- GRADE: Automatic Graph-Enhanced Coherence Metric for Evaluating Open-Domain Dialogue SystemsLishan Huang, Zheng Ye, Jinghui Qin, Liang Lin 等EMNLP 2020 · 被引用 73 次
- USR: An Unsupervised and Reference Free Evaluation Metric for Dialog GenerationShikib Mehri, Maxine EskénaziACL 2020 · 被引用 10 次
相关 Paper
- Semantic Representation for Dialogue ModelingXuefeng Bai, Yulong Chen, Linfeng Song, Yue ZhangACL 2021
- DEAM: Dialogue Coherence Evaluation using AMR-based Semantic ManipulationsSarik Ghazarian, Nuan Wen, Aram Galstyan, Nanyun PengACL 2022
- DynaEval: Unifying Turn and Dialogue Level EvaluationChen Zhang, Yiming Chen, Luis Fernando D'Haro, Yan Zhang 等ACL 2021
- Grounded Conversation Generation as Guided Traverses in Commonsense Knowledge GraphsHouyu Zhang, Zhenghao Liu, Chenyan Xiong, Zhiyuan LiuACL 2020 · 被引用 125 次
- Improving Knowledge-Aware Dialogue Generation via Knowledge Base Question AnsweringJian Wang, Junhao Liu, Wei Bi, Xiaojiang Liu 等AAAI 2020 · 被引用 52 次
