EvoGrad: Efficient Gradient-Based Meta-Learning and Hyperparameter Optimization
Ondrej Bohdal, Yongxin Yang, Timothy M. Hospedales
Abstract
Gradient-based meta-learning and hyperparameter optimization have seen significant progress recently, enabling practical end-to-end training of neural networks together with many hyperparameters. Nevertheless, existing approaches are relatively expensive as they need to compute second-order derivatives and store a longer computational graph. This cost prevents scaling them to larger network architectures. We present EvoGrad, a new approach to meta-learning that draws upon evolutionary techniques to more efficiently compute hypergradients. EvoGrad estimates hypergradient with respect to hyperparameters without calculating second-order gradients, or storing a longer computational graph, leading to significant improvements in efficiency. We evaluate EvoGrad on three substantial recent meta-learning applications, namely cross-domain few-shot learning with feature-wise transformations, noisy label learning with Meta-Weight-Net and low-resource cross-lingual learning with meta representation transformation. The results show that EvoGrad significantly improves efficiency and enables scaling meta-learning to bigger architectures such as from ResNet10 to ResNet34.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bd61e69e-548d-4610-a60d-75c8b3477c59Cited by top-tier papers5
- Bidirectional Learning for Offline Model-based Biological Sequence DesignCan Chen, Yingxue Zhang, Xue Liu, Mark CoatesICML 2023 · 30 citations
- MetaStore: Analyzing Deep Learning Meta-Data at ScaleHuayi Zhang, Binwei Yan, Lei Cao, Samuel Madden et al.VLDB 2024 · 10 citations
- The Pursuit of Human Labeling: A New Perspective on Unsupervised LearningArtyom Gadetsky, Maria BrbicNeurIPS 2023 · 10 citations
- Injecting Multimodal Information into Rigid Protein Docking via Bi-level OptimizationRuijia Wang, YiWu Sun, Yujie Luo, Shaochuan Li et al.NeurIPS 2023 · 9 citations
- Efficient Hyperparameter Optimization for LLM Reinforcement LearningMinping Chen, Bowen Xiao, Du Liang, Chuxuan Zeng et al.ACL 2026
Builds on2
Related papers
- MetaNorm: Learning to Normalize Few-Shot Batches Across DomainsYing-Jun Du, Xiantong Zhen, Ling Shao, Cees G. M. SnoekICLR 2021 · 26 citations
- Gradient-based Hyperparameter Optimization Over Long HorizonsPaul Micaelli, Amos J. StorkeyNeurIPS 2021 · 23 citations
- ES-MAML: Simple Hessian-Free Meta LearningXingyou Song, Wenbo Gao, Yuxiang Yang, Krzysztof Choromanski et al.ICLR 2020 · 128 citations
- Meta-Learning of Neural Architectures for Few-Shot LearningThomas Elsken, Benedikt Staffler, Jan Hendrik Metzen, Frank HutterCVPR 2020
- Meta-Learning via Learning with Distributed MemorySudarshan Babu, Pedro Savarese, Michael MaireNeurIPS 2021
