Meta-GMVAE: Mixture of Gaussian VAE for Unsupervised Meta-Learning
Dong Bok Lee, Dongchan Min, Seanie Lee, Sung Ju Hwang
Abstract
Unsupervised learning aims to learn meaningful representations from unlabeled data which can capture its intrinsic structure, that can be transferred to downstream tasks. Meta-learning, whose objective is to learn to generalize across tasks such that the learned model can rapidly adapt to a novel task, shares the spirit of unsupervised learning in that the both seek to learn more effective and efficient learning procedure than learning from scratch. The fundamental difference of the two is that the most meta-learning approaches are supervised, assuming full access to the labels. However, acquiring labeled dataset for meta-training not only is costly as it requires human efforts in labeling but also limits its applications to pre-defined task distributions. In this paper, we propose a principled unsupervised meta-learning model, namely Meta-GMVAE, based on Variational Autoencoder (VAE) and set-level variational inference. Moreover, we introduce a mixture of Gaussian (GMM) prior, assuming that each modality represents each class-concept in a randomly sampled episode, which we optimize with Expectation-Maximization (EM). Then, the learned model can be used for downstream few-shot classification tasks, where we obtain task-specific parameters by performing semi-supervised EM on the latent representations of the support and query set, and predict labels of the query set by computing aggregated posteriors. We validate our model on Omniglot and Mini-ImageNet datasets by evaluating its performance on downstream few-shot classification tasks. The results show that our model obtains impressive performance gains over existing unsupervised metalearning baselines, even outperforming supervised MAML on a certain setting.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers14
- Identifiability of deep generative models without auxiliary informationBohdan Kivva, Goutham Rajendran, Pradeep Ravikumar, Bryon AragamNeurIPS 2022 · 87 citations
- Shape your Space: A Gaussian Mixture Regularization Approach to Deterministic AutoencodersAmrutha Saseendran, Kathrin Skubch, Stefan Falkner, Margret KeuperNeurIPS 2021 · 13 citations
- CMVAE: Causal Meta VAE for Unsupervised Meta-LearningGuodong Qi, Huimin YuAAAI 2023 · 11 citations
- BECLR: Batch Enhanced Contrastive Few-Shot LearningStylianos Poulakakis-Daktylidis, Hadi Jamali RadICLR 2024 · 10 citations
- Big Learning Expectation MaximizationYulai Cong, Sijia LiAAAI 2024 · 5 citations
Builds on3
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Simple and Effective VAE Training with Calibrated DecodersOleh Rybkin, Kostas Daniilidis, Sergey LevineICML 2021 · 119 citations
- Guided Variational Autoencoder for Disentanglement LearningZheng Ding, Yifan Xu, Weijian Xu, Gaurav Parmar et al.CVPR 2020
Related papers
- Learning to Balance: Bayesian Meta-Learning for Imbalanced and Out-of-distribution TasksHaebeom Lee, Hayeon Lee, Donghyun Na, Saehoon Kim et al.ICLR 2020 · 115 citations
- Empirical Bayes Transductive Meta-Learning with Synthetic GradientsShell Xu Hu, Pablo Garcia Moreno, Yang Xiao, Xi Shen et al.ICLR 2020 · 139 citations
- Constructing Multiple Tasks for Augmentation: Improving Neural Image Classification with K-Means FeaturesTao Gui, Lizhi Qing, Qi Zhang, Jiacheng Ye et al.AAAI 2020 · 2 citations
- Addressing Catastrophic Forgetting in Few-Shot ProblemsPau Ching Yap, Hippolyt Ritter, David BarberICML 2021 · 20 citations
- Task Cooperation for Semi-Supervised Few-Shot LearningHan-Jia Ye, Xin-Chun Li, De-Chuan ZhanAAAI 2021 · 20 citations
