Autoregressive Conditional Neural Processes
Wessel P. Bruinsma, Stratis Markou, James Requeima, Andrew Y. K. Foong, Tom R. Andersson, Anna Vaughan, Anthony Buonomo, J. Scott Hosking, Richard E. Turner
摘要
Conditional neural processes (CNPs; Garnelo et al., 2018a) are attractive meta-learning models which produce well-calibrated predictions and are trainable via a simple maximum likelihood procedure. Although CNPs have many advantages, they are unable to model dependencies in their predictions. Various works propose solutions to this, but these come at the cost of either requiring approximate inference or being limited to Gaussian predictions. In this work, we instead propose to change how CNPs are deployed at test time, without any modifications to the model or training procedure. Instead of making predictions independently for every target point, we autoregressively define a joint predictive distribution using the chain rule of probability, taking inspiration from the neural autoregressive density estimator (NADE) literature. We show that this simple procedure allows factorised Gaussian CNPs to model highly dependent, non-Gaussian predictive distributions. Perhaps surprisingly, in an extensive range of tasks with synthetic and real data, we show that CNPs in autoregressive (AR) mode not only significantly outperform non-AR CNPs, but are also competitive with more sophisticated models that are significantly more computationally expensive and challenging to train. This performance is remarkable given that AR CNPs are not trained to model joint dependencies. Our work provides an example of how ideas from neural distribution estimation can benefit neural processes, and motivates research into the AR deployment of other neural process models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper24
- LLM Processes: Numerical Predictive Distributions Conditioned on Natural LanguageJames Requeima, John Bronskill, Dami Choi, Richard E. Turner 等NeurIPS 2024 · 被引用 72 次
- Neural Diffusion ProcessesVincent Dutordoir, Alan Saul, Zoubin Ghahramani, Fergus SimpsonICML 2023 · 被引用 52 次
- Amortized Bayesian Experimental Design for Decision-MakingDaolang Huang, Yujia Guo, Luigi Acerbi, Samuel KaskiNeurIPS 2024 · 被引用 24 次
- Episodic Multi-Task Learning with Heterogeneous Neural ProcessesJiayi Shen, Xiantong Zhen, Qi Wang, Marcel WorringNeurIPS 2023 · 被引用 21 次
- Practical Equivariances via Relational Conditional Neural ProcessesDaolang Huang, Manuel Haussmann, Ulpu Remes, S. T. John 等NeurIPS 2023 · 被引用 14 次
它引用的顶会 Paper6
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Convolutional Conditional Neural ProcessesJonathan Gordon, Wessel P. Bruinsma, Andrew Y. K. Foong, James Requeima 等ICLR 2020 · 被引用 200 次
- Transformer Neural Processes: Uncertainty-Aware Meta Learning Via Sequence ModelingTung Nguyen, Aditya GroverICML 2022 · 被引用 148 次
- Meta-Learning Stationary Stochastic Process Prediction with Convolutional Neural ProcessesAndrew Y. K. Foong, Wessel P. Bruinsma, Jonathan Gordon, Yann Dubois 等NeurIPS 2020 · 被引用 96 次
- Very Deep VAEs Generalize Autoregressive Models and Can Outperform Them on ImagesRewon ChildICLR 2021 · 被引用 45 次
相关 Paper
- Practical Conditional Neural Process Via Tractable Dependent PredictionsStratis Markou, James Requeima, Wessel P. Bruinsma, Anna Vaughan 等ICLR 2022 · 被引用 29 次
- Global Perception Based Autoregressive Neural ProcessesJinyang TaiICCV 2023 · 被引用 1 次
- NPCL: Neural Processes for Uncertainty-Aware Continual LearningSaurav Jha, Dong Gong, He Zhao, Lina YaoNeurIPS 2023 · 被引用 27 次
- Test Time Scaling for Neural ProcessesHyungi Lee, Moonseok Choi, Hyunsu Kim, Kyunghyun Cho 等NeurIPS 2025 · 被引用 1 次
- Contrastive Conditional Neural ProcessesZesheng Ye, Lina YaoCVPR 2022 · 被引用 11 次
