Enabling Few-Shot Learning with PID Control: A Layer Adaptive Optimizer
Le Yu, Xinde Li, Pengfei Zhang, Zhentong Zhang, Fir Dunkin
Abstract
Model-Agnostic Meta-Learning (MAML) and its variants have shown remarkable performance in scenarios characterized by a scarcity of labeled data during the training phase of machine learning models. Despite these successes, MAMLbased approaches encounter significant challenges when there is a substantial discrepancy in the distribution of training and testing tasks, resulting in inefficient learning and limited generalization across domains. Inspired by classical proportional-integral-derivative (PID) control theory, this study introduces a Layer-Adaptive PID (LA-PID) Optimizer, a MAML-based optimizer that employs efficient parameter optimization methods to dynamically adjust task-specific PID control gains at each layer of the network, conducting a first-principles analysis of optimal convergence conditions. A series of experiments conducted on four standard benchmark datasets demonstrate the efficacy of the LA-PID optimizer, indicating that LA-PID achieves state-ofthe-art performance in few-shot classification and cross-domain tasks, accomplishing these objectives with fewer training steps.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on7
- Rapid Learning or Feature Reuse? Towards Understanding the Effectiveness of MAMLAniruddh Raghu, Maithra Raghu, Samy Bengio, Oriol VinyalsICLR 2020 · 736 citations
- A Closer Look at Few-shot Classification AgainXu Luo, Hao Wu, Ji Zhang, Lianli Gao et al.ICML 2023 · 80 citations
- How to Train Your MAML to Excel in Few-Shot ClassificationHan-Jia Ye, Wei-Lun ChaoICLR 2022 · 61 citations
- SemSup-XC: Semantic Supervision for Zero and Few-shot Extreme ClassificationPranjal Aggarwal, Ameet Deshpande, Karthik R. NarasimhanICML 2023 · 8 citations
- Learning to Initialize: Can Meta Learning Improve Cross-task Generalization in Prompt Tuning?Chengwei Qin, Shafiq R. Joty, Qian Li, Ruochen ZhaoACL 2023 · 8 citations
Related papers
- Meta-Learning with a Geometry-Adaptive PreconditionerSuhyun Kang, Duhun Hwang, Moonjung Eo, Taesup Kim et al.CVPR 2023
- A Nested Bi-level Optimization Framework for Robust Few Shot LearningKrishnaTeja Killamsetty, Changbin Li, Chen Zhao, Feng Chen et al.AAAI 2022 · 12 citations
- Learning to Forget for Meta-LearningSungyong Baik, Seokil Hong, Kyoung Mu LeeCVPR 2020
- OOD-MAML: Meta-Learning for Few-Shot Out-of-Distribution Detection and ClassificationTaewon Jeong, Heeyoung KimNeurIPS 2020 · 111 citations
- Sharp-MAML: Sharpness-Aware Model-Agnostic Meta LearningMomin Abbas, Quan Xiao, Lisha Chen, Pin-Yu Chen et al.ICML 2022 · 105 citations
