DeliData: A Dataset for Deliberation in Multi-party Problem Solving
Georgi Karadzhov, Tom Stafford, Andreas Vlachos
Abstract
Group deliberation enables people to collaborate and solve problems, however, it is understudied due to a lack of resources. To this end, we introduce the first publicly available dataset containing collaborative conversations on solving a well-established cognitive task, consisting of 500 group dialogues and 14k utterances. In 64% of these conversations, the group members are able to find a better solution than they had identified individually, and in 43.8% of the groups who had a correct answer as their final solution, none of the participants had solved the task correctly by themselves. Furthermore, we propose a novel annotation schema that captures deliberation cues and release all 14k utterances annotated with it. Finally, we use the proposed dataset to develop and evaluate two methods for generating deliberation utterances. The data collection platform, dataset and annotated corpus are publicly available at https://delibot.xyz.
CCS Concepts: • Human-centered computing → Empirical studies in collaborative and social computing.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- Frictional Agent Alignment Framework: Slow Down and Don't Break ThingsAbhijnan Nath, Carine Graff, Andrei Bachinin, Nikhil KrishnaswamyACL 2025 · 8 citations
- Segment-Level Diffusion: A Framework for Controllable Long-Form Generation with Diffusion Language ModelsXiaochen Zhu, Georgi Karadzhov, Chenxi Whitehouse, Andreas VlachosACL 2025 · 3 citations
- Learning "Partner-Aware" Collaborators in Multi-Party CollaborationAbhijnan Nath, Nikhil KrishnaswamyNeurIPS 2025 · 2 citations
- Wisdom of the Crowd, Without the Crowd: A Socratic LLM for Asynchronous Deliberation on Perspectivist DataMalik Khadar, Daniel Runningen, Julia Tang, Stevie Chancellor et al.CSCW 2025 · 2 citations
- Evaluation and Facilitation of Online Discussions in the LLM Era: A SurveyKaterina Korre, Dimitris Tsirmpas, Nikos Gkoumas, Emma Cabalé et al.EMNLP 2025
Builds on5
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger et al.ICLR 2020 · 8,443 citations
- Large Dual Encoders Are Generalizable RetrieversJianmo Ni, Chen Qu, Jing Lu, Zhuyun Dai et al.EMNLP 2022 · 145 citations
- Moderator Chatbot for Deliberative Discussion: Effects of Discussion Structure and Discussant FacilitationSoomin Kim, Jinsu Eun, Joseph Seering, Joonhwan LeeCSCW 2021 · 84 citations
- What Changed Your Mind: The Roles of Dynamic Topics and Discourse in Argumentation ProcessJichuan Zeng, Jing Li, Yulan He, Cuiyun Gao et al.WWW 2020 · 16 citations
- Towards Argument Mining for Social Good: A SurveyEva Maria Vecchi, Neele Falk, Iman Jundi, Gabriella LapesaACL 2021
Related papers
- SuperDialseg: A Large-scale Dataset for Supervised Dialogue SegmentationJunfeng Jiang, Chengzhang Dong, Sadao Kurohashi, Akiko AizawaEMNLP 2023 · 2 citations
- MMDialog: A Large-scale Multi-turn Dialogue Dataset Towards Multi-modal Open-domain ConversationJiazhan Feng, Qingfeng Sun, Can Xu, Pu Zhao et al.ACL 2023 · 20 citations
- BotsTalk: Machine-sourced Framework for Automatic Curation of Large-scale Multi-skill Dialogue DatasetsMinju Kim, Chaehyeong Kim, Yongho Song, Seung-won Hwang et al.EMNLP 2022 · 8 citations
- ECFCON: Emotion Consequence Forecasting in ConversationsXincheng Ju, Dong Zhang, Suyang Zhu, Junhui Li et al.ACM MM 2024 · 1 citation
- YouRefIt: Embodied Reference Understanding with Language and GestureYixin Chen, Qing Li, Deqian Kong, Yik Lun Kei et al.ICCV 2021 · 57 citations
