Meta-Learning Priors Using Unrolled Proximal Networks
Yilang Zhang, Georgios B. Giannakis
Abstract
Relying on prior knowledge accumulated from related tasks, meta-learning offers a powerful approach to learning a novel task from limited training data. Recent approaches parameterize the prior with a family of probability density functions or recurrent neural networks, whose parameters can be optimized by utilizing validation data from the observed tasks. While these approaches have appealing empirical performance, the expressiveness of their prior is relatively low, which limits the generalization and interpretation of meta-learning. Aiming at expressive yet meaningful priors, this contribution puts forth a novel prior representation model that leverages the notion of algorithm unrolling. The key idea is to unroll the proximal gradient descent steps, where learnable piecewise linear functions are developed to approximate the desired proximal operators within tight theoretical error bounds established for both smooth and non-smooth proximal functions. The resultant multi-block neural network not only broadens the scope of learnable priors, but also enhances interpretability from an optimization viewpoint. Numerical tests conducted on few-shot learning datasets demonstrate markedly improved performance with flexible, visualizable, and understandable priors.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on10
- Meta-Learning with Warped Gradient DescentSebastian Flennerhag, Andrei A. Rusu, Razvan Pascanu, Francesco Visin et al.ICLR 2020 · 221 citations
- Meta-Learning with Task-Adaptive Loss Function for Few-Shot LearningSungyong Baik, Janghoon Choi, Heewon Kim, Dohee Cho et al.ICCV 2021 · 146 citations
- Proximal Denoiser for Convergent Plug-and-Play Optimization with Nonconvex RegularizationSamuel Hurault, Arthur Leclaire, Nicolas PapadakisICML 2022 · 121 citations
- Sharp-MAML: Sharpness-Aware Model-Agnostic Meta LearningMomin Abbas, Quan Xiao, Lisha Chen, Pin-Yu Chen et al.ICML 2022 · 105 citations
- Convergence of Meta-Learning with Task-Specific Adaptation over Partial ParametersKaiyi Ji, Jason D. Lee, Yingbin Liang, H. Vincent PoorNeurIPS 2020 · 97 citations
Related papers
- Meta-Learning Universal Priors Using Non-Injective Change of VariablesYilang Zhang, Alireza Sadeghi, Georgios B. GiannakisNeurIPS 2024 · 1 citation
- BOIL: Towards Representation Change for Few-shot LearningJaehoon Oh, Hyungjun Yoo, ChangHwan Kim, Se-Young YunICLR 2021 · 185 citations
- MARS: Meta-learning as Score Matching in the Function SpaceKrunoslav Lehman Pavasovic, Jonas Rothfuss, Andreas KrauseICLR 2023 · 1 citation
- MetaFun: Meta-Learning with Iterative Functional UpdatesJin Xu, Jean-Francois Ton, Hyunjik Kim, Adam R. Kosiorek et al.ICML 2020 · 76 citations
- Learning to Learn Dense Gaussian Processes for Few-Shot LearningZe Wang, Zichen Miao, Xiantong Zhen, Qiang QiuNeurIPS 2021 · 32 citations
