Enabling Few-Shot Learning with PID Control: A Layer Adaptive Optimizer
Le Yu, Xinde Li, Pengfei Zhang, Zhentong Zhang, Fir Dunkin
摘要
Model-Agnostic Meta-Learning (MAML) and its variants have shown remarkable performance in scenarios characterized by a scarcity of labeled data during the training phase of machine learning models. Despite these successes, MAMLbased approaches encounter significant challenges when there is a substantial discrepancy in the distribution of training and testing tasks, resulting in inefficient learning and limited generalization across domains. Inspired by classical proportional-integral-derivative (PID) control theory, this study introduces a Layer-Adaptive PID (LA-PID) Optimizer, a MAML-based optimizer that employs efficient parameter optimization methods to dynamically adjust task-specific PID control gains at each layer of the network, conducting a first-principles analysis of optimal convergence conditions. A series of experiments conducted on four standard benchmark datasets demonstrate the efficacy of the LA-PID optimizer, indicating that LA-PID achieves state-ofthe-art performance in few-shot classification and cross-domain tasks, accomplishing these objectives with fewer training steps.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper7
- Rapid Learning or Feature Reuse? Towards Understanding the Effectiveness of MAMLAniruddh Raghu, Maithra Raghu, Samy Bengio, Oriol VinyalsICLR 2020 · 被引用 736 次
- A Closer Look at Few-shot Classification AgainXu Luo, Hao Wu, Ji Zhang, Lianli Gao 等ICML 2023 · 被引用 80 次
- How to Train Your MAML to Excel in Few-Shot ClassificationHan-Jia Ye, Wei-Lun ChaoICLR 2022 · 被引用 61 次
- SemSup-XC: Semantic Supervision for Zero and Few-shot Extreme ClassificationPranjal Aggarwal, Ameet Deshpande, Karthik R. NarasimhanICML 2023 · 被引用 8 次
- Learning to Initialize: Can Meta Learning Improve Cross-task Generalization in Prompt Tuning?Chengwei Qin, Shafiq R. Joty, Qian Li, Ruochen ZhaoACL 2023 · 被引用 8 次
相关 Paper
- Meta-Learning with a Geometry-Adaptive PreconditionerSuhyun Kang, Duhun Hwang, Moonjung Eo, Taesup Kim 等CVPR 2023
- A Nested Bi-level Optimization Framework for Robust Few Shot LearningKrishnaTeja Killamsetty, Changbin Li, Chen Zhao, Feng Chen 等AAAI 2022 · 被引用 12 次
- Learning to Forget for Meta-LearningSungyong Baik, Seokil Hong, Kyoung Mu LeeCVPR 2020
- OOD-MAML: Meta-Learning for Few-Shot Out-of-Distribution Detection and ClassificationTaewon Jeong, Heeyoung KimNeurIPS 2020 · 被引用 111 次
- Sharp-MAML: Sharpness-Aware Model-Agnostic Meta LearningMomin Abbas, Quan Xiao, Lisha Chen, Pin-Yu Chen 等ICML 2022 · 被引用 105 次
