Efficient Allocation of Working Memory Resource for Utility Maximization in Humans and Recurrent Neural Networks
Qingqing Yang, Hsin-Hung Li
Abstract
Working memory (WM) supports the temporary retention of task-relevant information. It is limited in capacity and inherently noisy. The ability to flexibly allocate WM resource is a hallmark of adaptive behavior. While it is well established that WM resource can be prioritized via selective attention, whether they can be allocated based on reward incentive alone remains under debate—raising open questions about whether humans can efficiently allocate WM resource based on utility . To address this, we conducted behavioral experiments using orientations as stimuli. Participants first learned stimulus–reward associations and then performed a delayed estimate WM task. We found that WM precision, indexed by the variability of memory reports, reflected both natural stimulus priors and utility-based allocation. The effects from reward and prior on memory variability both grew over time, indicating their effects in stabilizing memory representations. In contrast, memory bias was largely unaffected by time or reward. To interpret these findings, we extended efficient coding theory by incorporating time and reformulating the objective from minimizing estimation loss to maximizing expected utility. We showed that the behavioral results were consistent with an observer that efficiently allocates WM resource over time to maximize utility. Lastly, we trained recurrent neural networks (RNNs) to perform the same WM task under a 2×2 design: prior (uniform vs. natural) × reward policy (baseline vs. reward context). Human-like behaviors emerged in RNNs: memory was more stable (lower variability) for stimuli associated with higher probability or rewards, and these effects increased over time. Transfer learning showed that recurrent dynamics were crucial for adapting to different priors and reward policies. Together, these results provide converging behavioral and computational evidence that WM resource allocation is shaped by environmental statistics and rewards, offering insight into how intelligent systems can dynamically optimize memory for utility under resource constraints.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 16c364a9-2ff3-45d6-9cc5-5e78b1828c26Related papers
- Dynamic allocation of limited memory resources in reinforcement learningNisheet Patel, Luigi Acerbi, Alexandre PougetNeurIPS 2020 · 6 citations
- Geometry of naturalistic object representations in recurrent neural network models of working memoryXiaoxuan Lei, Takuya Ito, Pouya BashivanNeurIPS 2024 · 1 citation
- Learning efficient task-dependent representations with synaptic plasticityColin Bredenberg, Eero P. Simoncelli, Cristina SavinNeurIPS 2020 · 10 citations
- Memory efficiency and resource-rational encoding in sentence processingWeijie Xu, Brian Dillon, Richard FutrellACL 2026
- RTify: Aligning Deep Neural Networks with Human Behavioral DecisionsYu-Ang Cheng, Ivan F. Rodriguez Rodriguez, Sixuan Chen, Kohitij Kar et al.NeurIPS 2024 · 11 citations
