Beyond Markovian Forgetfulness: Episodic Memory for Reasoning-Intensive Retrieval
Dohyeon Lee, Yeonseok Jeong, Seung-won Hwang
Abstract
Reasoning-intensive information retrieval uses large language models to solve complex queries via multi-step reasoning. However, existing methods have critical limitations. Chain-of-Thought (CoT) approaches suffer from inefficiency, while state-based methods, despite better token efficiency, often fall into reasoning cycles that trap the query refinement process. To address these issues, we propose Episodic Memory for Retrieval (EMR), which enhances the state-based framework with an episodic memory. This module stores the full history of prior states for a query, allowing the model to avoid repetition of such cycles. Experiments on the BRIGHT benchmark show that EMR consistently outperforms both CoT and state-based baselines. Moreover, it is highly token-efficient, reducing token usage by 72% on average. Our results show that episodic memory is an effective and tokenefficient mechanism for reasoning-intensive retrieval. The gains also generalize across different base models and stay efficient in terms of end-to-end latency. The code is available in https://github.com/ldilab/EMR .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on3
- RaDeR: Reasoning-aware Dense Retrieval ModelsDebrup Das, Seán Ó Nualláin, Razieh RahimiEMNLP 2025 · 1 citation
- ReAct: Synergizing Reasoning and Acting in Language ModelsShunyu Yao, Jeffrey Zhao, Dian Yu, Nan Du et al.ICLR 2023
- BRIGHT: A Realistic and Challenging Benchmark for Reasoning-Intensive RetrievalHongjin Su, Howard Yen, Mengzhou Xia, Weijia Shi et al.ICLR 2025
Related papers
- A State-Transition Framework for Efficient LLM ReasoningLiang Zhang, Yu Zhao, Longyue Wang, Tianqi Shi et al.ICLR 2026 · 2 citations
- Retrieval-of-Thought: Efficient Reasoning via Reusing ThoughtsAmmar Ahmed, Azal Ahmad Khan, Ayaan Ahmad, Sheng Di et al.ICLR 2026 · 13 citations
- Rethinking Reasoning in Document Ranking: Why Chain-of-Thought Falls ShortXuan Lu, Haohang Huang, Rui Meng, Yaohui Jin et al.ICLR 2026 · 11 citations
- ARTEM: Enhancing Large Language Model Agents with Spatial-Temporal Episodic MemoryCassandra Hui-Ming Tan, Budhitama Subagdja, Ah-Hwee TanAAAI 2026
- RAG without Forgetting: Continual Query-Infused Key MemoryYuntong Hu, Sha Li, Naren Ramakrishnan, Liang ZhaoICML 2026
