NaturalConv: A Chinese Dialogue Dataset Towards Multi-turn Topic-driven Conversation
Xiaoyang Wang, Chen Li, Jianqiao Zhao, Dong Yu
Abstract
In this paper, we propose a Chinese multi-turn topic-driven conversation dataset, NaturalConv, which allows the participants to chat anything they want as long as any element from the topic is mentioned and the topic shift is smooth. Our corpus contains 19.9K conversations from six domains, and 400K utterances with an average turn number of 20.1. These conversations contain in-depth discussions on related topics or widely natural transition between multiple topics. We believe either way is normal for human conversation. To facilitate the research on this corpus, we provide results of several benchmark models. Comparative results show that for this dataset, our current models are not able to provide significant improvement by introducing background knowledge/topic. Therefore, the proposed dataset should be a good benchmark for further research to evaluate the validity and naturalness of multi-turn conversation systems. Our dataset is available 1 at https://ailab.tencent.com/ailab/nlp/dialogue/#datasets .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9eb02a9e-08c3-41ec-9afb-ef9bc1797ca4Cited by top-tier papers9
- Where to Go for the Holidays: Towards Mixed-Type Dialogs for Clarification of User GoalsZeming Liu, Jun Xu, Zeyang Lei, Haifeng Wang et al.ACL 2022 · 18 citations
- Beyond Dialogue: A Profile-Dialogue Alignment Framework Towards General Role-Playing Language ModelYeyong Yu, Runsheng Yu, Haojie Wei, Zhanqiu Zhang et al.ACL 2025 · 13 citations
- CHBias: Bias Evaluation and Mitigation of Chinese Conversational Language ModelsJiaxu Zhao, Meng Fang, Zijing Shi, Yitong Li et al.ACL 2023 · 11 citations
- MOA: Multi-Objective Alignment for Role-Playing AgentsChonghua Liao, Ke Wang, Yuchuan Wu, Ruoran Li et al.ACL 2026 · 4 citations
- McHirc: A Multimodal Benchmark for Chinese Idiom Reading ComprehensionTongguan Wang, Mingmin Wu, Guixin Su, Dongyu Su et al.AAAI 2025 · 4 citations
Builds on2
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger et al.ICLR 2020 · 8,443 citations
- KdConv: A Chinese Multi-domain Dialogue Dataset Towards Multi-turn Knowledge-driven ConversationHao Zhou, Chujie Zheng, Kaili Huang, Minlie Huang et al.ACL 2020 · 106 citations
Related papers
- RealTalk-CN: A Realistic Chinese Speech Task-Oriented Dialogue Benchmark with Cross-Modal AnalysisEnzhi Wang, Jiaming Zhou, Yuhang Jia, Aobo Kong et al.ACL 2026
- MP2D: An Automated Topic Shift Dialogue Generation Framework Leveraging Knowledge GraphsYerin Hwang, Yongil Kim, Yunah Jang, Jeesoo Bang et al.EMNLP 2024 · 2 citations
- MMConv: An Environment for Multimodal Conversational Search across Multiple DomainsLizi Liao, Le Hong Long, Zheng Zhang, Minlie Huang et al.SIGIR 2021 · 70 citations
- CDConv: A Benchmark for Contradiction Detection in Chinese ConversationsChujie Zheng, Jinfeng Zhou, Yinhe Zheng, Libiao Peng et al.EMNLP 2022 · 6 citations
- CGoDial: A Large-Scale Benchmark for Chinese Goal-oriented Dialog EvaluationYinpei Dai, Wanwei He, Bowen Li, Yuchuan Wu et al.EMNLP 2022 · 6 citations
