CDConv: A Benchmark for Contradiction Detection in Chinese Conversations
Chujie Zheng, Jinfeng Zhou, Yinhe Zheng, Libiao Peng, Zhen Guo, Wenquan Wu, Zheng-Yu Niu, Hua Wu, Minlie Huang
Abstract
Dialogue contradiction is a critical issue in open-domain dialogue systems. The contextualization nature of conversations makes dialogue contradiction detection rather challenging. In this work, we propose a benchmark for Contradiction Detection in Chinese Conversations, namely CDCONV. It contains 12K multi-turn conversations annotated with three typical contradiction categories: Intrasentence Contradiction, Role Confusion, and History Contradiction. To efficiently construct the CDCONV conversations, we devise a series of methods for automatic conversation generation, which simulate common user behaviors that trigger chatbots to make contradictions. We conduct careful manual quality screening of the constructed conversations and show that state-of-the-art Chinese chatbots can be easily goaded into making contradictions. Experiments on CDCONV show that properly modeling contextual information is critical for dialogue contradiction detection, but there are still unresolved challenges that require future research. 1 * Equal contribution.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 91be4a7a-d47f-480c-8959-b5859f4c6a8eCited by top-tier papers4
- Knowledge Conflicts for LLMs: A SurveyRongwu Xu, Zehan Qi, Zhijiang Guo, Cunxiang Wang et al.EMNLP 2024 · 38 citations
- Exploring Conversational Adaptability: Assessing the Proficiency of Large Language Models in Dynamic Alignment with Updated User IntentYu-Chuan Chen, Hen-Hsen HuangAAAI 2025 · 2 citations
- Red Teaming Language Models for Processing Contradictory DialoguesXiaofei Wen, Bangzheng Li, Tenghao Huang, Muhao ChenEMNLP 2024 · 1 citation
- Think Wider, Detect Sharper: Reinforced Reference Coverage for Document-Level Self-Contradiction DetectionYuhao Chen, Yuanjie Lyu, Shuochen Liu, Chao Zhang et al.EMNLP 2025
Builds on7
- PLATO: Pre-trained Dialogue Generation Model with Discrete Latent VariableSiqi Bao, Huang He, Fan Wang, Hua Wu et al.ACL 2020 · 229 citations
- The Stem Cell Hypothesis: Dilemma behind Multi-Task Learning with Transformer EncodersHan He, Jinho D. ChoiEMNLP 2021 · 111 citations
- Dialog Inpainting: Turning Documents into DialogsZhuyun Dai, Arun Tejasvi Chaganty, Vincent Y. Zhao, Aida Amini et al.ICML 2022 · 77 citations
- Neural Path Hunter: Reducing Hallucination in Dialogue Systems via Path GroundingNouha Dziri, Andrea Madotto, Osmar Zaïane, Avishek Joey BoseEMNLP 2021 · 74 citations
- Profile Consistency Identification for Open-domain Dialogue AgentsHaoyu Song, Yan Wang, Wei-Nan Zhang, Zhengyu Zhao et al.EMNLP 2020 · 19 citations
Related papers
- I like fish, especially dolphins: Addressing Contradictions in Dialogue ModelingYixin Nie, Mary Williamson, Mohit Bansal, Douwe Kiela et al.ACL 2021
- CGoDial: A Large-Scale Benchmark for Chinese Goal-oriented Dialog EvaluationYinpei Dai, Wanwei He, Bowen Li, Yuchuan Wu et al.EMNLP 2022 · 6 citations
- CORECODE: A Common Sense Annotated Dialogue Dataset with Benchmark Tasks for Chinese Large Language ModelsDan Shi, Chaobin You, Jiantao Huang, Taihao Li et al.AAAI 2024 · 3 citations
- NaturalConv: A Chinese Dialogue Dataset Towards Multi-turn Topic-driven ConversationXiaoyang Wang, Chen Li, Jianqiao Zhao, Dong YuAAAI 2021 · 54 citations
- CSDS: A Fine-Grained Chinese Dataset for Customer Service Dialogue SummarizationHaitao Lin, Liqun Ma, Junnan Zhu, Lu Xiang et al.EMNLP 2021 · 21 citations
