Neural Mixture Density Processes
Yi Ding, Qi Tao, Xingxing Liang, Longfei Zhang, Yiqin Lv, Weitao Song, Fangjie Yang, Cheems Wang, Guangquan Cheng
摘要
The neural process (NP) is a probabilistic meta-learning model that learns distributions over functions via a global latent variable. It enables fast adaptation in few-shot scenarios by leveraging past experience. However, the design of latent variable structures and conditioning mechanisms in NPs remains underexplored, despite their importance in capturing diverse functional distributions. This paper proposes a new variant of NPs via mixture density modeling, referred to as the neural mixture density process (NMDP). The NMDP decomposes model parameters into task-agnostic and task-specific components to represent function distributions more flexibly. We train the model using a variational EM/MM-style procedure with self-normalized importance sampling, yielding an explicit surrogate objective for learning expressive functional priors. Compared with existing work, our method maintains several advantages: (i) efficient adaptation at test time by only inferring a compact taskspecific latent variable, (ii) compact task representation via distributions in the simplex, (iii) a principled EM/MM-style optimization with a monotonic-improvement guarantee in the idealized exact-inference setting. Experimental results show that our method can achieve competitive performance with adequate explainability.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper22
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Convolutional Conditional Neural ProcessesJonathan Gordon, Wessel P. Bruinsma, Andrew Y. K. Foong, James Requeima 等ICLR 2020 · 被引用 200 次
- Bayesian Meta-Learning for the Few-Shot Setting via Deep KernelsMassimiliano Patacchiola, Jack Turner, Elliot J. Crowley, Michael F. P. O'Boyle 等NeurIPS 2020 · 被引用 167 次
- Transformer Neural Processes: Uncertainty-Aware Meta Learning Via Sequence ModelingTung Nguyen, Aditya GroverICML 2022 · 被引用 148 次
- On Episodes, Prototypical Networks, and Few-Shot LearningSteinar Laenen, Luca BertinettoNeurIPS 2021 · 被引用 142 次
相关 Paper
- Learning Expressive Meta-Representations with Mixture of Expert Neural ProcessesQi Wang, Herke van HoofNeurIPS 2022 · 被引用 35 次
- Neural Variational Dropout ProcessesInsu Jeon, Youngjin Park, Gunhee KimICLR 2022 · 被引用 3 次
- Accurate Bayesian Meta-Learning by Accurate Task Posterior InferenceMichael Volpp, Philipp Dahlinger, Philipp Becker, Christian Daniel 等ICLR 2023
- Neural Diffusion ProcessesVincent Dutordoir, Alan Saul, Zoubin Ghahramani, Fergus SimpsonICML 2023 · 被引用 52 次
- MARS: Meta-learning as Score Matching in the Function SpaceKrunoslav Lehman Pavasovic, Jonas Rothfuss, Andreas KrauseICLR 2023 · 被引用 1 次
