Enhanced Story Comprehension for Large Language Models through Dynamic Document-Based Knowledge Graphs
Berkeley R. Andrus, Yeganeh Nasiri, Shilong Cui, Benjamin Cullen, Nancy Fulda
Abstract
Large transformer-based language models have achieved incredible success at various tasks which require narrative comprehension, including story completion, answering questions about stories, and generating stories ex nihilo. However, due to the limitations of finite context windows, these language models struggle to produce or understand stories longer than several thousand tokens. In order to mitigate the document length limitations that come with finite context windows, we introduce a novel architecture that augments story processing with an external dynamic knowledge graph. In contrast to static commonsense knowledge graphs which hold information about the real world, these dynamic knowledge graphs reflect facts extracted from the story being processed. Our architecture uses these knowledge graphs to create informationrich prompts which better facilitate story comprehension than prompts composed only of story text. We apply our architecture to the tasks of question answering and story completion. To complement this line of research, we introduce two long-form question answering tasks, LF-SQuAD and LF-QUOREF, in which the document length exceeds the size of the language model's context window, and introduce a story completion evaluation method that bypasses the stochastic nature of language model generation. We demonstrate broad improvement over typical prompt formulation methods for both question answering and story completion using GPT-2, GPT-3 and XLNet.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0a7bbc9d-d8c0-4eee-9665-8fb1690b2fc5Cited by top-tier papers6
- Prompting Large Language Models with Chain-of-Thought for Few-Shot Knowledge Base Question GenerationYuanyuan Liang, Jianing Wang, Hanlun Zhu, Lei Wang et al.EMNLP 2023 · 24 citations
- Embodied CoT Distillation From LLM To Off-the-shelf AgentsWonje Choi, Woo Kyung Kim, Minjong Yoo, Honguk WooICML 2024 · 13 citations
- Exploring the Design Space of Real-time LLM Knowledge Support Systems: A Case Study of Jargon ExplanationsYuhan Liu, Aadit Shah, Jordan Ackerman, Manaswi SahaCHI 2025 · 9 citations
- Understand the Dynamic World: An End-to-End Knowledge Informed Framework for Open Domain Entity State TrackingMingchen Li, Lifu HuangSIGIR 2023 · 6 citations
- Synergizing Multimodal Temporal Knowledge Graphs and Large Language Models for Social Relation RecognitionHaorui Wang, Zheng Wang, Yuxuan Zhang, Bo Wang et al.EMNLP 2025 · 1 citation
Builds on6
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Retrieval Augmented Language Model Pre-TrainingKelvin Guu, Kenton Lee, Zora Tung, Panupong Pasupat et al.ICML 2020 · 2,937 citations
- Effective Modeling of Encoder-Decoder Architecture for Joint Entity and Relation ExtractionTapas Nayak, Hwee Tou NgAAAI 2020 · 272 citations
- ETC: Encoding Long and Structured Inputs in TransformersJoshua Ainslie, Santiago Ontañón, Chris Alberti, Vaclav Cvicek et al.EMNLP 2020 · 268 citations
- Dynamic Neuro-Symbolic Knowledge Graph Construction for Zero-shot Commonsense Question AnsweringAntoine Bosselut, Ronan Le Bras, Yejin ChoiAAAI 2021 · 135 citations
Related papers
- Learning Event Graph Knowledge for Abductive ReasoningLi Du, Xiao Ding, Ting Liu, Bing QinACL 2021
- Multiple Knowledge Syncretic Transformer for Natural Dialogue GenerationXiangyu Zhao, Longbiao Wang, Ruifang He, Ting Yang et al.WWW 2020 · 27 citations
- A Human-Inspired Reading Agent with Gist Memory of Very Long ContextsKuang-Huei Lee, Xinyun Chen, Hiroki Furuta, John F. Canny et al.ICML 2024 · 106 citations
- Fine-Grained Modeling of Narrative Context: A Coherence Perspective via Retrospective QuestionsLiyan Xu, Jiangnan Li, Mo Yu, Jie ZhouACL 2024
- Evaluating Commonsense in Pre-Trained Language ModelsXuhui Zhou, Yue Zhang, Leyang Cui, Dandan HuangAAAI 2020 · 198 citations
