Symbolic Working Memory Enhances Language Models for Complex Rule Application
Siyuan Wang, Zhongyu Wei, Yejin Choi, Xiang Ren
Abstract
Large Language Models (LLMs) have shown remarkable reasoning performance but struggle with multi-step deductive reasoning involving a series of rule application steps, especially when rules are presented non-sequentially. Our preliminary analysis shows that while LLMs excel in single-step rule application, their performance drops significantly in multi-step scenarios due to the challenge in rule grounding. It requires anchoring the applicable rule and supporting facts at each step, amidst multiple input rules, facts, and inferred facts. To address this, we propose augmenting LLMs with external working memory and introduce a neurosymbolic framework for rule application. The memory stores facts and rules in both natural language and symbolic forms, enabling precise tracking. Utilizing this memory, our framework iteratively performs symbolic rule grounding and LLM-based rule implementation. The former matches predicates and variables of symbolic rules and facts to ground applicable rules at each step. Experiments indicate our framework's effectiveness in rule application and its robustness across various steps and settings 1 . 1 Code and data are available at https://github.com/ SiyuanWangw/RuleApplication . [Sequential Input] Facts: Nicole's grandfather, Harold, accompanied her to the basketball match. (F1) Beverly went car shopping with her husband Louis and her daughter Nicole. (F2) Harold bought a new dress for his daughter Marie. (F3) Rules: If B is A's daughter, and C is B's grandfather, then C is the father of A. (R1) If B is the father of A, and C is the daughter of B, then C is the sister of A. (R2) [Non-Sequential Input] Facts: Harold bought a new dress for his daughter Marie. (F3) Nicole's grandfather, Harold, accompanied her to the basketball match. (F1) Beverly went car shopping with her husband Louis and her daughter Nicole. (F2) Rules: If B is A's father, and C is B's daughter, then C is the sister of A. (R2) If B is A's daughter, and C is B's grandfather, then C is the father of A. (R1) [Query] How is Marie related to Beverly? [Rule Application Order]: R1→ (F2+F1) ⟹ F4; R2 → (F4+F3) ⟹ Answer
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c5d10ac7-c8bd-476f-9370-ca63d814002eCited by top-tier papers6
- RuleArena: A Benchmark for Rule-Guided Reasoning with LLMs in Real-World ScenariosRuiwen Zhou, Wenyue Hua, Liangming Pan, Sitao Cheng et al.ACL 2025 · 13 citations
- RuleReasoner: Reinforced Rule-based Reasoning via Domain-aware Dynamic SamplingYang Liu, Jiaqi Li, Zilong ZhengICLR 2026 · 8 citations
- MAGNET: Towards Adaptive GUI Agents with Memory-Driven Knowledge EvolutionLibo Sun, Jiwen Zhang, Siyuan Wang, Zhongyu WeiACL 2026 · 5 citations
- Leibniz: Theory-of-Mind Driven Neuro-Symbolic Logical Reasoning via Multi-Agent CollaborationYue Fan, Hu Zhang, Yunxiao Zhao, Guangjun Zhang et al.ACL 2026
- Can Compressed LLMs Truly Act? An Empirical Evaluation of Agentic Capabilities in LLM CompressionPeijie Dong, Zhenheng Tang, Xiang Liu, Lujun Li et al.ICML 2025
Builds on14
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- Generative Agents: Interactive Simulacra of Human BehaviorJoon Sung Park, Joseph C. O'Brien, Carrie Jun Cai, Meredith Ringel Morris et al.UIST 2023 · 1,882 citations
- Self-Consistency Improves Chain of Thought Reasoning in Language ModelsXuezhi Wang, Jason Wei, Dale Schuurmans, Quoc V. Le et al.ICLR 2023 · 681 citations
- Augmenting Language Models with Long-Term MemoryWeizhi Wang, Li Dong, Hao Cheng, Xiaodong Liu et al.NeurIPS 2023 · 256 citations
- Phenomenal Yet Puzzling: Testing Inductive Reasoning Capabilities of Language Models with Hypothesis RefinementLinlu Qiu, Liwei Jiang, Ximing Lu, Melanie Sclar et al.ICLR 2024 · 114 citations
Related papers
- Advancing Multimodal Agent Reasoning with Long-Term Neuro-Symbolic MemoryRongjie Jiang, Jianwei Wang, Gengda Zhao, Chengyang Luo et al.KDD 2026 · 5 citations
- RLIE: Rule Generation with Logistic Regression, Iterative Refinement, and Evaluation for Large Language ModelsYang Yang, Hua XU, Zhangyi Hu, Yutao YueICML 2026
- Logically Consistent Language Models via Neuro-Symbolic IntegrationDiego Calanzone, Stefano Teso, Antonio VergariICLR 2025 · 2 citations
- Faithful Logical Reasoning via Symbolic Chain-of-ThoughtJundong Xu, Hao Fei, Liangming Pan, Qian Liu et al.ACL 2024
- RIMRULE: Improving Tool-Using Language Agents via MDL-Guided Rule LearningXiang Gao, Yuguang Yao, Qi Zhang, Kaiwen Dong et al.ACL 2026 · 2 citations
