SHOT-VAE: Semi-supervised Deep Generative Models With Label-aware ELBO Approximations
Haozhe Feng, Kezhi Kong, Minghao Chen, Tianye Zhang, Minfeng Zhu, Wei Chen
摘要
Semi-supervised variational autoencoders (VAEs) have obtained strong results, but have also encountered the challenge that good ELBO values do not always imply accurate inference results. In this paper, we investigate and propose two causes of this problem: (1) The ELBO objective cannot utilize the label information directly. (2) A bottleneck value exists, and continuing to optimize ELBO after this value will not improve inference accuracy. On the basis of the experiment results, we propose SHOT-VAE to address these problems without introducing additional prior knowledge. The SHOT-VAE offers two contributions: (1) A new ELBO approximation named smooth-ELBO that integrates the label predictive loss into ELBO. (2) An approximation based on optimal interpolation that breaks the ELBO value bottleneck by reducing the margin between ELBO and the data likelihood. The SHOT-VAE achieves good performance with 25.30% error rate on CIFAR-100 with 10k labels and reduces the error rate to 6.11% on CIFAR-10 with 4k labels.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- CauseRec: Counterfactual User Sequence Synthesis for Sequential RecommendationShengyu Zhang, Dong Yao, Zhou Zhao, Tat-Seng Chua 等SIGIR 2021 · 被引用 118 次
- Class-Incremental Instance Segmentation via Multi-Teacher NetworksYanan Gu, Cheng Deng, Kun WeiAAAI 2021 · 被引用 32 次
- Invariant Action Effect Model for Reinforcement LearningZheng-Mao Zhu, Shengyi Jiang, Yu-Ren Liu, Yang Yu 等AAAI 2022 · 被引用 12 次
- StrWAEs to Invariant RepresentationsHyunjong Lee, Yedarm Seong, Sungdong Lee, Joong-Ho WonICML 2024
它引用的顶会 Paper2
相关 Paper
- VAE Approximation Error: ELBO and Exponential FamiliesAlexander Shekhovtsov, Dmitrij Schlesinger, Boris FlachICLR 2022 · 被引用 21 次
- PG-LBO: Enhancing High-Dimensional Bayesian Optimization with Pseudo-Label and Gaussian Process GuidanceTaicai Chen, Yue Duan, Dong Li, Lei Qi 等AAAI 2024 · 被引用 12 次
- Learning Mask Invariant Mutual Information for Masked Image ModelingTao Huang, Yanxiang Ma, Shan You, Chang XuICLR 2025
- Capturing Label Characteristics in VAEsTom Joy, Sebastian M. Schmon, Philip H. S. Torr, Siddharth Narayanaswamy 等ICLR 2021 · 被引用 54 次
- On the Limitations of Multimodal VAEsImant Daunhawer, Thomas M. Sutter, Kieran Chin-Cheong, Emanuele Palumbo 等ICLR 2022 · 被引用 50 次
