ORCHID: A Chinese Debate Corpus for Target-Independent Stance Detection and Argumentative Dialogue Summarization
Xiutian Zhao, Ke Wang, Wei Peng
摘要
Dialogue agents have been receiving increasing attention for years, and this trend has been further boosted by the recent progress of large language models (LLMs). Stance detection and dialogue summarization are two core tasks of dialogue agents in application scenarios that involve argumentative dialogues. However, research on these tasks is limited by the insufficiency of public datasets, especially for non-English languages. To address this language resource gap in Chinese, we present OR-CHID (Oral Chinese Debate), the first Chinese dataset for benchmarking target-independent stance detection and debate summarization. Our dataset consists of 1,218 real-world debates that were conducted in Chinese on 476 unique topics, containing 2,436 stance-specific summaries and 14,133 fully annotated utterances. Besides providing a versatile testbed for future research, we also conduct an empirical study on the dataset and propose an integrated task. The results show the challenging nature of the dataset and suggest a potential of incorporating stance detection in summarization for argumentative dialogue. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper8
- DialogLM: Pre-trained Model for Long Dialogue Understanding and SummarizationMing Zhong, Yang Liu, Yichong Xu, Chenguang Zhu 等AAAI 2022 · 被引用 150 次
- Topic-Oriented Spoken Dialogue Summarization for Customer Service with Saliency-Aware Topic ModelingYicheng Zou, Lujun Zhao, Yangyang Kang, Jun Lin 等AAAI 2021 · 被引用 63 次
- CSDS: A Fine-Grained Chinese Dataset for Customer Service Dialogue SummarizationHaitao Lin, Liqun Ma, Junnan Zhu, Lu Xiang 等EMNLP 2021 · 被引用 21 次
- Zero-Shot Stance Detection: A Dataset and Model using Generalized Topic RepresentationsEmily Allaway, Kathleen R. McKeownEMNLP 2020 · 被引用 7 次
- IAM: A Comprehensive and Large-Scale Dataset for Integrated Argument Mining TasksLiying Cheng, Lidong Bing, Ruidan He, Qian Yu 等ACL 2022
相关 Paper
- C-STANCE: A Large Dataset for Chinese Zero-Shot Stance DetectionChenye Zhao, Yingjie Li, Cornelia CarageaACL 2023 · 被引用 12 次
- Towards Multi-dimensional Evaluation of LLM Summarization across Domains and LanguagesHyangsuk Min, Yuho Lee, Minjeong Ban, Jiaqi Deng 等ACL 2025 · 被引用 8 次
- Diversity Over Size: On the Effect of Sample and Topic Sizes for Topic-Dependent Argument Mining DatasetsBenjamin Schiller, Johannes Daxenberger, Andreas Waldis, Iryna GurevychEMNLP 2024 · 被引用 3 次
- Bilingual Zero-Shot Stance DetectionChenye Zhao, Cornelia CarageaACL 2025 · 被引用 1 次
- Towards Understanding Omission in Dialogue SummarizationYicheng Zou, Kaitao Song, Xu Tan, Zhongkai Fu 等ACL 2023 · 被引用 3 次
