Injecting Context via Situation Working Memory for Logical Reasoning with LLMs
Jieun Kim, Seoha Lim, YoungHae Choi, Sung-Bae Cho
摘要
Recent advances in large language models (LLMs) have improved logical reasoning by incorporating formal logic or explicit structured representations. However, such methods often lose track of what is true now in multistep reasoning, failing to maintain a coherent global state and its logical consequences. Motivated by Situation Model Theory in cognitive psychology, which views comprehension as constructing and updating a mental model of events along key dimensions (time, space, causality, intention, protagonist), we propose a cognitively inspired method of Situation Working Memory (SituW) for contextual reasoning in LLMs. SituW first builds a situation representation by decomposing text along these five dimensions, and guides LLM inference with the evolving state. Keeping an explicit, dynamically updated situation memory instead of a static logical form encourages globally consistent reasoning over the situation model rather than raw text. Evaluated in both supervised and prompt-based settings, SituW improves accuracy by 23.3%p and 15.93%p while reducing "uncertain" predictions, suggesting that explicit situation modeling supports more globally consistent LLM reasoning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper13
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- PAL: Program-aided Language ModelsLuyu Gao, Aman Madaan, Shuyan Zhou, Uri Alon 等ICML 2023 · 被引用 700 次
- SatLM: Satisfiability-Aided Language Models Using Declarative PromptingXi Ye, Qiaochu Chen, Isil Dillig, Greg DurrettNeurIPS 2023 · 被引用 126 次
- Neuro-Symbolic Inductive Logic Programming with Logical Neural NetworksPrithviraj Sen, Breno W. S. R. de Carvalho, Ryan Riegel, Alexander G. GrayAAAI 2022 · 被引用 82 次
相关 Paper
- Structure Guided Prompt: Instructing Large Language Model in Multi-Step Reasoning by Exploring Graph Structure of the TextKewei Cheng, Nesreen K. Ahmed, Theodore L. Willke, Yizhou SunEMNLP 2024 · 被引用 6 次
- HGMem: Hypergraph-based Working Memory to Improve Multi-step RAG for Long-Context Complex Relational ModelingChulun Zhou, Chunkang Zhang, Guoxin Yu, Fandong Meng 等ICML 2026 · 被引用 4 次
- Selection-Inference: Exploiting Large Language Models for Interpretable Logical ReasoningAntonia Creswell, Murray Shanahan, Irina HigginsICLR 2023 · 被引用 110 次
- Hypothesis-Driven Reasoning for Large Language ModelsAakash Kumar Agarwal, Moyuru YamadaAAAI 2026
- Learning to Reason and Memorize with Self-NotesJack Lanchantin, Shubham Toshniwal, Jason Weston, Arthur Szlam 等NeurIPS 2023 · 被引用 45 次
