Exploring Stable Meta-Optimization Patterns via Differentiable Reinforcement Learning for Few-Shot Classification
Zheng Han, Xiaobin Zhu, Chun Yang, Hongyang Zhou, Jingyan Qin, Xu-Cheng Yin
摘要
Existing few-shot learning methods generally focus on designing exquisite structures of meta-learners for learning task-specific prior to improve the discriminative ability of global embeddings. However, they often ignore the importance of learning stability in meta-training, making it difficult to obtain a relatively optimal model. From this key observation, we propose an innovative generic differentiable Reinforcement Learning (RL) strategy for few-shot classification. It aims to explore stable meta-optimization patterns in meta-training by learning generalizable optimizations for producing task-adaptive embeddings. Accordingly, our differentiable RL strategy models the embedding procedure of feature transformation layers in meta-learner to optimize the gradient flow implicitly. Also, we propose a memory module to associate historical and current task states and actions for exploring inter-task similarity. Notably, our RL-based strategy can be easily extended to various backbones. In addition, we propose a novel task state encoder to encode task representation, which fully explores inner-task similarities between support set and query set. Extensive experiments verify that our approach can improve the performance of different backbones and achieve promising results against state-of-the-art methods in few-shot classification.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- Reinforced Attention for Few-Shot Learning and BeyondJie Hong, Pengfei Fang, Weihao Li, Tong Zhang 等CVPR 2021
- CAD: Co-Adapting Discriminative Features for Improved Few-Shot ClassificationPhilip Chikontwe, Soopil Kim, Sang Hyun ParkCVPR 2022 · 被引用 46 次
- Few-Shot Video Classification via Representation Fusion and Promotion LearningHaifeng Xia, Kai Li, Martin Renqiang Min, Zhengming DingICCV 2023 · 被引用 13 次
- A Closer Look at the Training Strategy for Modern Meta-LearningJiaxin Chen, Xiao-Ming Wu, Yanke Li, Qimai Li 等NeurIPS 2020 · 被引用 48 次
- DVLA-RL: Dual-Level Vision-Language Alignment with Reinforcement Learning Gating for Few-Shot LearningWenhao Li, Xianjing Meng, Qiangchang Wang, Zhongyi Han 等ICLR 2026 · 被引用 4 次
