COINS: Dynamically Generating COntextualized Inference Rules for Narrative Story Completion
Debjit Paul, Anette Frank
Abstract
Despite recent successes of large pre-trained language models in solving reasoning tasks, their inference capabilities remain opaque. We posit that such models can be made more interpretable by explicitly generating interim inference rules, and using them to guide the generation of task-specific textual outputs. In this paper we present COINS, a recursive inference framework that i) iteratively reads context sentences, ii) dynamically generates contextualized inference rules, encodes them, and iii) uses them to guide task-specific output generation. We apply COINS to a Narrative Story Completion task that asks a model to complete a story with missing sentences, to produce a coherent story with plausible logical connections, causal relationships, and temporal dependencies. By modularizing inference and sentence generation steps in a recurrent model, we aim to make reasoning steps and their effects on next sentence generation transparent. Our automatic and manual evaluations show that the model generates better story sentences than SOTA baselines, especially in terms of coherence. We further demonstrate improved performance over strong pre-trained LMs in generating commonsense inference rules. The recursive nature of COINS holds the potential for controlled generation of longer sequences.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on9
- Abductive Commonsense ReasoningChandra Bhagavatula, Ronan Le Bras, Chaitanya Malaviya, Keisuke Sakaguchi et al.ICLR 2020 · 521 citations
- ASER: A Large-scale Eventuality Knowledge GraphHongming Zhang, Xin Liu, Haojie Pan, Yangqiu Song et al.WWW 2020 · 183 citations
- Commonsense Knowledge Base Completion with Structural and Semantic ContextChaitanya Malaviya, Chandra Bhagavatula, Antoine Bosselut, Yejin ChoiAAAI 2020 · 155 citations
- MEGATRON-CNTRL: Controllable Story Generation with External Knowledge Using Large-Scale Language ModelsPeng Xu, Mostofa Patwary, Mohammad Shoeybi, Raul Puri et al.EMNLP 2020 · 104 citations
- Language Generation with Multi-Hop Reasoning on Commonsense Knowledge GraphHaozhe Ji, Pei Ke, Shaohan Huang, Furu Wei et al.EMNLP 2020 · 96 citations
Related papers
- Selection-Inference: Exploiting Large Language Models for Interpretable Logical ReasoningAntonia Creswell, Murray Shanahan, Irina HigginsICLR 2023 · 110 citations
- Language Models of Code are Few-Shot Commonsense LearnersAman Madaan, Shuyan Zhou, Uri Alon, Yiming Yang et al.EMNLP 2022 · 103 citations
- SCOUT: Teaching Pre-trained Language Models to Enhance Reasoning via Flow Chain-of-ThoughtGuanghao Li, Wenhao Jiang, Mingfeng Chen, Yan Li et al.NeurIPS 2025 · 8 citations
- Towards Interpretable Reasoning over Paragraph Effects in SituationMucheng Ren, Xiubo Geng, Tao Qin, Heyan Huang et al.EMNLP 2020 · 6 citations
- Nash CoT: Multi-Path Inference with Preference EquilibriumZiqi Zhang, Cunxiang Wang, Xiao Xiong, Yue Zhang et al.EMNLP 2024 · 1 citation
