Paragraph-level Commonsense Transformers with Recurrent Memory
Saadia Gabriel, Chandra Bhagavatula, Vered Shwartz, Ronan Le Bras, Maxwell Forbes, Yejin Choi
Abstract
Human understanding of narrative texts requires making commonsense inferences beyond what is stated explicitly in the text. A recent model, COMET, can generate such implicit commonsense inferences along several dimensions such as pre-and post-conditions, motivations, and mental states of the participants. However, COMET was trained on commonsense inferences of short phrases, and is therefore discourseagnostic. When presented with each sentence of a multisentence narrative, it might generate inferences that are inconsistent with the rest of the narrative. We present the task of discourse-aware commonsense inference. Given a sentence within a narrative, the goal is to generate commonsense inferences along predefined dimensions, while maintaining coherence with the rest of the narrative. Such large-scale paragraph-level annotation is hard to get and costly, so we use available sentence-level annotations to efficiently and automatically construct a distantly supervised corpus. Using this corpus, we train PARA-COMET, a discourseaware model that incorporates paragraph-level information to generate coherent commonsense inferences from narratives. PARA-COMET captures both semantic knowledge pertaining to prior world knowledge, and episodic knowledge involving how current events relate to prior and future events in a narrative. Our results show that PARA-COMET outperforms the sentence-level baselines, particularly in generating inferences that are both coherent and novel.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b34f6371-0a50-4e68-9475-244b687ffaecCited by top-tier papers5
- ValueNet: A New Dataset for Human Value Driven Dialogue SystemLiang Qiu, Yizhou Zhao, Jinchao Li, Pan Lu et al.AAAI 2022 · 51 citations
- Entailer: Answering Questions with Faithful and Truthful Chains of ReasoningOyvind Tafjord, Bhavana Dalvi Mishra, Peter ClarkEMNLP 2022 · 28 citations
- Using Commonsense Knowledge to Answer Why-QuestionsYash Kumar Lal, Niket Tandon, Tanvi Aggarwal, Horace Liu et al.EMNLP 2022 · 7 citations
- Complex Reasoning over Logical Queries on Commonsense Knowledge GraphsTianqing Fang, Zeming Chen, Yangqiu Song, Antoine BosselutACL 2024 · 5 citations
- COINS: Dynamically Generating COntextualized Inference Rules for Narrative Story CompletionDebjit Paul, Anette FrankACL 2021
Builds on3
- Automated Storytelling via Causal, Commonsense Plot OrderingPrithviraj Ammanabrolu, Wesley Cheung, William Broniec, Mark O. RiedlAAAI 2021 · 72 citations
- R^3: Reverse, Retrieve, and Rank for Sarcasm Generation with Commonsense KnowledgeTuhin Chakrabarty, Debanjan Ghosh, Smaranda Muresan, Nanyun PengACL 2020 · 58 citations
- Social Bias Frames: Reasoning about Social and Power Implications of LanguageMaarten Sap, Saadia Gabriel, Lianhui Qin, Dan Jurafsky et al.ACL 2020 · 16 citations
Related papers
- EventBERT: A Pre-Trained Model for Event Correlation ReasoningYucheng Zhou, Xiubo Geng, Tao Shen, Guodong Long et al.WWW 2022 · 66 citations
- DiffuCOMET: Contextual Commonsense Knowledge DiffusionSilin Gao, Mete Ismayilzada, Mengjie Zhao, Hiromi Wakaki et al.ACL 2024 · 2 citations
- DISCOS: Bridging the Gap between Discourse Knowledge and Commonsense KnowledgeTianqing Fang, Hongming Zhang, Weiqi Wang, Yangqiu Song et al.WWW 2021 · 48 citations
- Pretraining with Contrastive Sentence Objectives Improves Discourse Performance of Language ModelsDan Iter, Kelvin Guu, Larry Lansing, Dan JurafskyACL 2020 · 72 citations
- Modeling Human Motives and Emotions from Personal Narratives Using External Knowledge And Entity TrackingPrashanth Vijayaraghavan, Deb RoyWWW 2021 · 10 citations
