Continuous-Time Meta-Learning with Forward Mode Differentiation
Tristan Deleu, David Kanaa, Leo Feng, Giancarlo Kerg, Yoshua Bengio, Guillaume Lajoie, Pierre-Luc Bacon
摘要
Drawing inspiration from gradient-based meta-learning methods with infinitely small gradient steps, we introduce Continuous-Time Meta-Learning (COMLN), a meta-learning algorithm where adaptation follows the dynamics of a gradient vector field. Specifically, representations of the inputs are meta-learned such that a task-specific linear classifier is obtained as a solution of an ordinary differential equation (ODE). Treating the learning process as an ODE offers the notable advantage that the length of the trajectory is now continuous, as opposed to a fixed and discrete number of gradient steps. As a consequence, we can optimize the amount of adaptation necessary to solve a new task using stochastic gradient descent, in addition to learning the initial conditions as is standard practice in gradient-based meta-learning. Importantly, in order to compute the exact meta-gradients required for the outer-loop updates, we devise an efficient algorithm based on forward mode differentiation, whose memory requirements do not scale with the length of the learning trajectory, thus allowing longer adaptation in constant memory. We provide analytical guarantees for the stability of COMLN, we show empirically its efficiency in terms of runtime and memory usage, and we illustrate its effectiveness on a range of few-shot image classification problems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- MetaDiff: Meta-Learning with Conditional Diffusion for Few-Shot LearningBaoquan Zhang, Chuyao Luo, Demin Yu, Xutao Li 等AAAI 2024 · 被引用 91 次
- Class-Aware Patch Embedding Adaptation for Few-Shot Image ClassificationFusheng Hao, Fengxiang He, Liu Liu, Fuxiang Wu 等ICCV 2023 · 被引用 56 次
- Influencing Long-Term Behavior in Multiagent Reinforcement LearningDong-Ki Kim, Matthew Riemer, Miao Liu, Jakob N. Foerster 等NeurIPS 2022 · 被引用 29 次
- Neural Differential Equations for Learning to Program Neural Nets Through Continuous Learning RulesKazuki Irie, Francesco Faccio, Jürgen SchmidhuberNeurIPS 2022 · 被引用 24 次
- Revisiting Neural Networks for Few-Shot Learning: A Zero-Cost NAS PerspectiveHaidong KangICML 2025
它引用的顶会 Paper8
- Gradient Surgery for Multi-Task LearningTianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine 等NeurIPS 2020 · 被引用 2,261 次
- Rapid Learning or Feature Reuse? Towards Understanding the Effectiveness of MAMLAniruddh Raghu, Maithra Raghu, Samy Bengio, Oriol VinyalsICLR 2020 · 被引用 736 次
- Meta-Learning with Warped Gradient DescentSebastian Flennerhag, Andrei A. Rusu, Razvan Pascanu, Francesco Visin 等ICLR 2020 · 被引用 221 次
- BOIL: Towards Representation Change for Few-shot LearningJaehoon Oh, Hyungjun Yoo, ChangHwan Kim, Se-Young YunICLR 2021 · 被引用 185 次
- Learning where to learn: Gradient sparsity in meta and continual learningJohannes von Oswald, Dominic Zhao, Seijin Kobayashi, Simon Schug 等NeurIPS 2021 · 被引用 61 次
相关 Paper
- On Enforcing Better Conditioned Meta-Learning for Rapid Few-Shot AdaptationMarkus Hiller, Mehrtash Harandi, Tom DrummondNeurIPS 2022 · 被引用 10 次
- Gradient-based Hyperparameter Optimization Over Long HorizonsPaul Micaelli, Amos J. StorkeyNeurIPS 2021 · 被引用 23 次
- MetaNODE: Prototype Optimization as a Neural ODE for Few-Shot LearningBaoquan Zhang, Xutao Li, Shanshan Feng, Yunming Ye 等AAAI 2022 · 被引用 46 次
- MetaFun: Meta-Learning with Iterative Functional UpdatesJin Xu, Jean-Francois Ton, Hyunjik Kim, Adam R. Kosiorek 等ICML 2020 · 被引用 76 次
- A Lazy Approach to Long-Horizon Gradient-Based Meta-LearningMuhammad Abdullah Jamal, Liqiang Wang, Boqing GongICCV 2021 · 被引用 9 次
