Enhanced Story Comprehension for Large Language Models through Dynamic Document-Based Knowledge Graphs
Berkeley R. Andrus, Yeganeh Nasiri, Shilong Cui, Benjamin Cullen, Nancy Fulda
摘要
Large transformer-based language models have achieved incredible success at various tasks which require narrative comprehension, including story completion, answering questions about stories, and generating stories ex nihilo. However, due to the limitations of finite context windows, these language models struggle to produce or understand stories longer than several thousand tokens. In order to mitigate the document length limitations that come with finite context windows, we introduce a novel architecture that augments story processing with an external dynamic knowledge graph. In contrast to static commonsense knowledge graphs which hold information about the real world, these dynamic knowledge graphs reflect facts extracted from the story being processed. Our architecture uses these knowledge graphs to create informationrich prompts which better facilitate story comprehension than prompts composed only of story text. We apply our architecture to the tasks of question answering and story completion. To complement this line of research, we introduce two long-form question answering tasks, LF-SQuAD and LF-QUOREF, in which the document length exceeds the size of the language model's context window, and introduce a story completion evaluation method that bypasses the stochastic nature of language model generation. We demonstrate broad improvement over typical prompt formulation methods for both question answering and story completion using GPT-2, GPT-3 and XLNet.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Prompting Large Language Models with Chain-of-Thought for Few-Shot Knowledge Base Question GenerationYuanyuan Liang, Jianing Wang, Hanlun Zhu, Lei Wang 等EMNLP 2023 · 被引用 24 次
- Embodied CoT Distillation From LLM To Off-the-shelf AgentsWonje Choi, Woo Kyung Kim, Minjong Yoo, Honguk WooICML 2024 · 被引用 13 次
- Exploring the Design Space of Real-time LLM Knowledge Support Systems: A Case Study of Jargon ExplanationsYuhan Liu, Aadit Shah, Jordan Ackerman, Manaswi SahaCHI 2025 · 被引用 9 次
- Understand the Dynamic World: An End-to-End Knowledge Informed Framework for Open Domain Entity State TrackingMingchen Li, Lifu HuangSIGIR 2023 · 被引用 6 次
- Synergizing Multimodal Temporal Knowledge Graphs and Large Language Models for Social Relation RecognitionHaorui Wang, Zheng Wang, Yuxuan Zhang, Bo Wang 等EMNLP 2025 · 被引用 1 次
它引用的顶会 Paper6
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Retrieval Augmented Language Model Pre-TrainingKelvin Guu, Kenton Lee, Zora Tung, Panupong Pasupat 等ICML 2020 · 被引用 2,937 次
- Effective Modeling of Encoder-Decoder Architecture for Joint Entity and Relation ExtractionTapas Nayak, Hwee Tou NgAAAI 2020 · 被引用 272 次
- ETC: Encoding Long and Structured Inputs in TransformersJoshua Ainslie, Santiago Ontañón, Chris Alberti, Vaclav Cvicek 等EMNLP 2020 · 被引用 268 次
- Dynamic Neuro-Symbolic Knowledge Graph Construction for Zero-shot Commonsense Question AnsweringAntoine Bosselut, Ronan Le Bras, Yejin ChoiAAAI 2021 · 被引用 135 次
相关 Paper
- Learning Event Graph Knowledge for Abductive ReasoningLi Du, Xiao Ding, Ting Liu, Bing QinACL 2021
- Multiple Knowledge Syncretic Transformer for Natural Dialogue GenerationXiangyu Zhao, Longbiao Wang, Ruifang He, Ting Yang 等WWW 2020 · 被引用 27 次
- A Human-Inspired Reading Agent with Gist Memory of Very Long ContextsKuang-Huei Lee, Xinyun Chen, Hiroki Furuta, John F. Canny 等ICML 2024 · 被引用 106 次
- Fine-Grained Modeling of Narrative Context: A Coherence Perspective via Retrospective QuestionsLiyan Xu, Jiangnan Li, Mo Yu, Jie ZhouACL 2024
- Evaluating Commonsense in Pre-Trained Language ModelsXuhui Zhou, Yue Zhang, Leyang Cui, Dandan HuangAAAI 2020 · 被引用 198 次
