CICERO: A Dataset for Contextualized Commonsense Inference in Dialogues
Deepanway Ghosal, Siqi Shen, Navonil Majumder, Rada Mihalcea, Soujanya Poria
Abstract
This paper addresses the problem of dialogue reasoning with contextualized commonsense inference. We curate CICERO, a dataset of dyadic conversations with five types of utterance-level reasoning-based inferences: cause, subsequent event, prerequisite, motivation, and emotional reaction. The dataset contains 53,105 of such inferences from 5,672 dialogues. We use this dataset to solve relevant generative and discriminative tasks: generation of cause and subsequent event; generation of prerequisite, motivation, and listener's emotional reaction; and selection of plausible alternatives. Our results ascertain the value of such dialogue-centric commonsense knowledge datasets. It is our hope that CI-CERO will open new research avenues into commonsense-based dialogue reasoning. A: Can I help you? B: Yes, please. I'd like some oranges. A: Do you want Florida or California oranges? B: Which do you think are better? A: Florida oranges are sweet but they are small. But California oranges have no seeds. B: Then give me five California oranges. A: Anything else? B: I also want some bananas. How do you sell them? A: One dollar a pound. How many do you want? B: Give me four and see how much they are.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers11
- ArCHer: Training Language Model Agents via Hierarchical Multi-Turn RLYifei Zhou, Andrea Zanette, Jiayi Pan, Sergey Levine et al.ICML 2024 · 163 citations
- Dialogue Chain-of-Thought Distillation for Commonsense-aware Conversational AgentsHyungjoo Chae, Yongho Song, Kai Tzu-iunn Ong, Taeyoon Kwon et al.EMNLP 2023 · 16 citations
- Reverse Multi-Choice Dialogue Commonsense Inference with Graph-of-ThoughtLi Zheng, Hao Fei, Fei Li, Bobo Li et al.AAAI 2024 · 13 citations
- CRoW: Benchmarking Commonsense Reasoning in Real-World TasksMete Ismayilzada, Debjit Paul, Syrielle Montariol, Mor Geva et al.EMNLP 2023 · 3 citations
- A Textual Dataset for Situated Proactive Response SelectionNaoki Otani, Jun Araki, HyeongSik Kim, Eduard H. HovyACL 2023 · 1 citation
Builds on5
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 2,496 citations
- Abductive Commonsense ReasoningChandra Bhagavatula, Ronan Le Bras, Chaitanya Malaviya, Keisuke Sakaguchi et al.ICLR 2020 · 521 citations
- (Comet-) Atomic 2020: On Symbolic and Neural Commonsense Knowledge GraphsJena D. Hwang, Chandra Bhagavatula, Ronan Le Bras, Jeff Da et al.AAAI 2021 · 458 citations
- MuTual: A Dataset for Multi-Turn Dialogue ReasoningLeyang Cui, Yu Wu, Shujie Liu, Yue Zhang et al.ACL 2020 · 115 citations
- GLUCOSE: GeneraLized and COntextualized Story ExplanationsNasrin Mostafazadeh, Aditya Kalyanpur, Lori Moon, David W. Buchanan et al.EMNLP 2020 · 7 citations
Related papers
- CORECODE: A Common Sense Annotated Dialogue Dataset with Benchmark Tasks for Chinese Large Language ModelsDan Shi, Chaobin You, Jiantao Huang, Taihao Li et al.AAAI 2024 · 3 citations
- PCoKG: Personality-aware Commonsense Reasoning with DebateWeijie Li, Zhongqing Wang, Guodong ZhouAAAI 2026
- CELLO: Causal Evaluation of Large Vision-Language ModelsMeiqi Chen, Bo Peng, Yan Zhang, Chaochao LuEMNLP 2024 · 4 citations
- ECC: An Emotion-Cause Conversation Dataset for Empathy ResponseYuanyuan He, Yongsen Pan, Wei Li, Jiali You et al.EMNLP 2025
- ACCENT: An Automatic Event Commonsense Evaluation Metric for Open-Domain Dialogue SystemsSarik Ghazarian, Yijia Shao, Rujun Han, Aram Galstyan et al.ACL 2023 · 3 citations
