Memory Efficient Neural Processes via Constant Memory Attention Block
Leo Feng, Frederick Tung, Hossein Hajimirsadeghi, Yoshua Bengio, Mohamed Osama Ahmed
Abstract
Neural Processes (NPs) are popular meta-learning methods for efficiently modelling predictive uncertainty. Recent state-of-the-art methods, however, leverage expensive attention mechanisms, limiting their applications, particularly in low-resource settings. In this work, we propose Constant Memory Attentive Neural Processes (CMANPs), an NP variant that only requires constant memory. To do so, we first propose an efficient update operation for Cross Attention. Leveraging the update operation, we propose Constant Memory Attention Block (CMAB), a novel attention block that (i) is permutation invariant, (ii) computes its output in constant memory, and (iii) performs constant computation updates. Finally, building on CMAB, we detail Constant Memory Attentive Neural Processes. Empirically, we show CMANPs achieve state-of-the-art results on popular NP benchmarks while being significantly more memory efficient than prior methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0bcbea93-60c3-4f87-8b25-7cdf71a22343Cited by top-tier papers6
- ALINE: Joint Amortization for Bayesian Inference and Active Data AcquisitionDaolang Huang, Xinyi Wen, Ayush Bharti, Samuel Kaski et al.NeurIPS 2025 · 8 citations
- Efficient Autoregressive Inference for Transformer Probabilistic ModelsConor Hassan, Nasrulloh R. B. S. Loka, Cen-You Li, Daolang Huang et al.ICLR 2026 · 5 citations
- Revisiting Synthetic Human Trajectories: Imitative Generation and Benchmarks Beyond DatasaurusBangchao Deng, Xin Jing, Tianyue Yang, Bingqing Qu et al.KDD 2025 · 3 citations
- Test Time Scaling for Neural ProcessesHyungi Lee, Moonseok Choi, Hyunsu Kim, Kyunghyun Cho et al.NeurIPS 2025 · 1 citation
- Gridded Transformer Neural Processes for Spatio-Temporal DataMatthew Ashman, Cristiana Diaconu, Eric Langezaal, Adrian Weller et al.ICML 2025
Builds on11
- Perceiver IO: A General Architecture for Structured Inputs & OutputsAndrew Jaegle, Sebastian Borgeaud, Jean-Baptiste Alayrac, Carl Doersch et al.ICLR 2022 · 797 citations
- Transformer Hawkes ProcessSimiao Zuo, Haoming Jiang, Zichong Li, Tuo Zhao et al.ICML 2020 · 382 citations
- Intensity-Free Learning of Temporal Point ProcessesOleksandr Shchur, Marin Bilos, Stephan GünnemannICLR 2020 · 210 citations
- Convolutional Conditional Neural ProcessesJonathan Gordon, Wessel P. Bruinsma, Andrew Y. K. Foong, James Requeima et al.ICLR 2020 · 200 citations
- Transformer Neural Processes: Uncertainty-Aware Meta Learning Via Sequence ModelingTung Nguyen, Aditya GroverICML 2022 · 148 citations
Related papers
- Latent Bottlenecked Attentive Neural ProcessesLeo Feng, Hossein Hajimirsadeghi, Yoshua Bengio, Mohamed Osama AhmedICLR 2023
- NPCL: Neural Processes for Uncertainty-Aware Continual LearningSaurav Jha, Dong Gong, He Zhao, Lina YaoNeurIPS 2023 · 27 citations
- Practical Equivariances via Relational Conditional Neural ProcessesDaolang Huang, Manuel Haussmann, Ulpu Remes, S. T. John et al.NeurIPS 2023 · 14 citations
- Dimension Agnostic Neural ProcessesHyungi Lee, Chaeyun Jang, Dongbok Lee, Juho LeeICLR 2025
- Robustifying Sequential Neural ProcessesJaesik Yoon, Gautam Singh, Sungjin AhnICML 2020 · 26 citations
