Gaussian Mixture Variational Autoencoder with Contrastive Learning for Multi-Label Classification
Junwen Bai, Shufeng Kong, Carla P. Gomes
Abstract
Multi-label classification (MLC) is a prediction task where each sample can have more than one label. We propose a novel contrastive learning boosted multi-label prediction model based on a Gaussian mixture variational autoencoder (C-GMVAE), which learns a multimodal prior space and employs a contrastive loss. Many existing methods introduce extra complex neural modules like graph neural networks to capture the label correlations, in addition to the prediction modules. We find that by using contrastive learning in the supervised setting, we can exploit label information effectively in a data-driven manner, and learn meaningful feature and label embeddings which capture the label correlations and enhance the predictive power. Our method also adopts the idea of learning and aligning latent spaces for both features and labels. In contrast to previous works based on a unimodal prior, C-GMVAE imposes a Gaussian mixture structure on the latent space, to alleviate the posterior collapse and over-regularization issues. C-GMVAE outperforms existing methods on multiple public datasets and can often match other models' full performance with only 50% of the training data. Furthermore, we show that the learnt embeddings provide insights into the interpretation of label-label interactions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9a7182c7-c8d3-4e60-aaf8-155a93dbfdd5Cited by top-tier papers15
- Multi-Label Supervised Contrastive LearningPingyue Zhang, Mengyue WuAAAI 2024 · 42 citations
- PatchCT: Aligning Patch Set and Label Set with Conditional Transport for Multi-Label Image ClassificationMiaoge Li, Dongsheng Wang, Xinyang Liu, Zequn Zeng et al.ICCV 2023 · 30 citations
- Label-Specific Feature Augmentation for Long-Tailed Multi-Label Text ClassificationPengyu Xu, Lin Xiao, Bing Liu, Sijin Lu et al.AAAI 2023 · 29 citations
- Partial Multi-View Multi-Label Classification via Semantic Invariance Learning and Prototype ModelingChengliang Liu, Gehui Xu, Jie Wen, Yabo Liu et al.ICML 2024 · 18 citations
- A Learnable Discrete-Prior Fusion Autoencoder with Contrastive Learning for Tabular Data SynthesisRongchao Zhang, Yiwei Lou, Dexuan Xu, Yongzhi Cao et al.AAAI 2024 · 14 citations
Builds on7
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna et al.NeurIPS 2020 · 7,049 citations
- Asymmetric Loss For Multi-Label ClassificationTal Ridnik, Emanuel Ben Baruch, Nadav Zamir, Asaf Noy et al.ICCV 2021 · 778 citations
- Multi-label Contrastive Predictive CodingJiaming Song, Stefano ErmonNeurIPS 2020 · 53 citations
- Relating by Contrasting: A Data-efficient Framework for Multimodal Generative ModelsYuge Shi, Brooks Paige, Philip H. S. Torr, N. SiddharthICLR 2021 · 42 citations
Related papers
- Multimodal Gaussian Mixture Variational Autoencoder with Consistency RegularizationsYarui Chen, Lehan Hong, Jianlin Shao, Jianning Yang et al.AAAI 2026
- Collaborative Graph Convolutional Networks: Unsupervised Learning Meets Semi-Supervised LearningBinyuan Hui, Pengfei Zhu, Qinghua HuAAAI 2020 · 67 citations
- Deep Multiview Clustering by Contrasting Cluster AssignmentsJie Chen, Hua Mao, Wai Lok Woo, Xi PengICCV 2023 · 142 citations
- M²VAE: Multi-Modal Multi-View Variational Autoencoder for Cold-start Item RecommendationChuan He, Yongchao Liu, Qiang Li, Chuntao Hong et al.AAAI 2026 · 1 citation
- Incomplete Multi-View Multi-label Learning via Disentangled Representation and Label Semantic EmbeddingXu Yan, Jun Yin, Jie WenCVPR 2025
