Hierarchical Gaussian Mixture based Task Generative Model for Robust Meta-Learning
Yizhou Zhang, Jingchao Ni, Wei Cheng, Zhengzhang Chen, Liang Tong, Haifeng Chen, Yan Liu
Abstract
Meta-learning enables quick adaptation of machine learning models to new tasks with limited data. While tasks could come from varying distributions in reality, most of the existing meta-learning methods consider both training and testing tasks as from the same uni-component distribution, overlooking two critical needs of a practical solution: (1) the various sources of tasks may compose a multi-component mixture distribution, and (2) novel tasks may come from a distribution that is unseen during meta-training. In this paper, we demonstrate these two challenges can be solved jointly by modeling the density of task instances. We develop a meta-training framework underlain by a novel Hierarchical Gaussian Mixture based Task Generative Model (HTGM). HTGM extends the widely used empirical process of sampling tasks to a theoretical model, which learns task embeddings, fits the mixture distribution of tasks, and enables density-based scoring of novel tasks. The framework is agnostic to the encoder and scales well with large backbone networks. The model parameters are learned end-to-end by maximum likelihood estimation via an Expectation-Maximization (EM) algorithm. Extensive experiments on benchmark datasets indicate the effectiveness of our method for both sample classification and novel task detection.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e9da7818-2bfe-468f-9de7-96e7932bd7f7Cited by top-tier papers1
Ask how each one uses itBuilds on19
- Energy-based Out-of-distribution DetectionWeitang Liu, Xiaoyun Wang, John D. Owens, Yixuan LiNeurIPS 2020 · 2,213 citations
- Rapid Learning or Feature Reuse? Towards Understanding the Effectiveness of MAMLAniruddh Raghu, Maithra Raghu, Samy Bengio, Oriol VinyalsICLR 2020 · 736 citations
- Open-Set Recognition: A Good Closed-Set Classifier is All You NeedSagar Vaze, Kai Han, Andrea Vedaldi, Andrew ZissermanICLR 2022 · 594 citations
- Task2Vec: Task Embedding for Meta-LearningAlessandro Achille, Michael Lam, Rahul Tewari, Avinash Ravichandran et al.ICCV 2019 · 359 citations
- On Episodes, Prototypical Networks, and Few-Shot LearningSteinar Laenen, Luca BertinettoNeurIPS 2021 · 142 citations
Related papers
- Meta-GMVAE: Mixture of Gaussian VAE for Unsupervised Meta-LearningDong Bok Lee, Dongchan Min, Seanie Lee, Sung Ju HwangICLR 2021 · 62 citations
- Secure Out-of-Distribution Task Generalization with Energy-Based ModelsShengzhuang Chen, Long-Kai Huang, Jonathan Richard Schwarz, Yilun Du et al.NeurIPS 2023 · 10 citations
- Task Aligned Generative Meta-learning for Zero-shot LearningZhe Liu, Yun Li, Lina Yao, Xianzhi Wang et al.AAAI 2021 · 45 citations
- Open Domain Generalization with Domain-Augmented Meta-LearningYang Shu, Zhangjie Cao, Chenyu Wang, Jianmin Wang et al.CVPR 2021
- Superclass-Conditional Gaussian Mixture Model For Learning Fine-Grained EmbeddingsJingchao Ni, Wei Cheng, Zhengzhang Chen, Takayoshi Asakura et al.ICLR 2022 · 15 citations
