A Dataset of Argumentative Dialogues on Scientific Papers
Federico Ruggeri, Mohsen Mesgar, Iryna Gurevych
Abstract
With recent advances in question-answering models, various datasets have been collected to improve and study the effectiveness of these models on scientific texts. Questions and answers in these datasets explore a scientific paper by seeking factual information from the paper's content. However, these datasets do not tackle the argumentative content of scientific papers, which is of huge importance in persuasiveness of a scientific discussion. We introduce ArgSciChat, a dataset of 41 argumentative dialogues between scientists on 20 NLP papers. The unique property of our dataset is that it includes both exploratory and argumentative questions and answers in a dialogue discourse on a scientific paper. Moreover, the size of ArgSciChat demonstrates the difficulties in collecting dialogues for specialized domains. Thus, our dataset is a challenging resource to evaluate dialogue agents in low-resource domains, in which collecting training data is costly. We annotate all sentences of dialogues in ArgSciChat and analyze them extensively. The results confirm that dialogues in ArgSci-Chat include exploratory and argumentative interactions. Furthermore, we use our dataset to fine-tune and evaluate a pre-trained documentgrounded dialogue agent. The agent achieves a low performance on our dataset, motivating a need for dialogue agents with a capability to reason and argue about their answers. We publicly release ArgSciChat 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 536fb633-d73e-4151-8be4-fb31e2317f91Cited by top-tier papers7
- LLMs Assist NLP Researchers: Critique Paper (Meta-)ReviewingJiangshu Du, Yibo Wang, Wenting Zhao, Zhongfen Deng et al.EMNLP 2024 · 14 citations
- PaperTrail: A Claim-Evidence Interface for Grounding Provenance in LLM-based Scholarly Q&AAnna Martin-Boyle, Cara A. C. Leckey, Martha Brown, Harmanpreet KaurCHI 2026 · 3 citations
- An Expert Schema for Evaluating Large Language Model Errors in Scholarly Question-Answering SystemsAnna Martin-Boyle, William Humphreys, Martha Brown, Cara A. C. Leckey et al.CHI 2026 · 1 citation
- SciMDR: Advancing Scientific Multimodal Document ReasoningZiyu Chen, Yilun Zhao, Chengye Wang, Rilyn Han et al.ACL 2026 · 1 citation
- Tackling the Root of Misinformation by Teaching Laypeople about Logical Fallacies via Socratic Questioning and Critical ArgumentationMinjing Shi, Junling Wang, Jingwei Ni, Sankalan Pal Chowdhury et al.ACL 2026
Builds on3
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger et al.ICLR 2020 · 8,443 citations
- doc2dial: A Goal-Oriented Document-Grounded Dialogue DatasetSong Feng, Hui Wan, R. Chulaka Gunasekara, Siva Sankalp Patel et al.EMNLP 2020 · 87 citations
- APE: Argument Pair Extraction from Peer Review and Rebuttal via Multi-task LearningLiying Cheng, Lidong Bing, Qian Yu, Wei Lu et al.EMNLP 2020 · 56 citations
Related papers
- QAConv: Question Answering on Informative ConversationsChien-Sheng Wu, Andrea Madotto, Wenhao Liu, Pascale Fung et al.ACL 2022 · 34 citations
- ArgueTutor: An Adaptive Dialog-Based Learning System for Argumentation SkillsThiemo Wambsganss, Tobias Kueng, Matthias Söllner, Jan Marco LeimeisterCHI 2021 · 126 citations
- SAD: A Large-Scale Strategic Argumentative Dialogue DatasetYongkang Liu, Jiayang Yu, Mingyang Wang, Yiqun Zhang et al.ACL 2026
- Language Models as Science TutorsAlexis Chevalier, Jiayi Geng, Alexander Wettig, Howard Chen et al.ICML 2024 · 17 citations
- Building and Evaluating Open-Domain Dialogue Corpora with Clarifying QuestionsMohammad Aliannejadi, Julia Kiseleva, Aleksandr Chuklin, Jeff Dalton et al.EMNLP 2021 · 61 citations
