Deep Generative Modeling on Limited Data with Regularization by Nontransferable Pre-trained Models
Yong Zhong, Hongtao Liu, Xiaodong Liu, Fan Bao, Weiran Shen, Chongxuan Li
Abstract
Deep generative models (DGMs) are data-eager because learning a complex model on limited data suffers from a large variance and easily overfits. Inspired by the classical perspective of the bias-variance tradeoff, we propose regularized deep generative model (Reg-DGM), which leverages a nontransferable pre-trained model to reduce the variance of generative modeling with limited data. Formally, Reg-DGM optimizes a weighted sum of a certain divergence and the expectation of an energy function, where the divergence is between the data and the model distributions, and the energy function is defined by the pre-trained model w.r.t. the model distribution. We analyze a simple yet representative Gaussian-fitting case to demonstrate how the weighting hyperparameter trades off the bias and the variance. Theoretically, we characterize the existence and the uniqueness of the global minimum of Reg-DGM in a non-parametric setting and prove its convergence with neural networks trained by gradient-based methods. Empirically, with various pre-trained feature extractors and a data-dependent energy function, Reg-DGM consistently improves the generation performance of strong DGMs with limited data and achieves competitive results to the state-of-the-art methods. Our implementation is available at https://github.com/ML-GSAI/Reg-ADA-APA.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Contrastive Energy Prediction for Exact Energy-Guided Diffusion Sampling in Offline Reinforcement LearningCheng Lu, Huayu Chen, Jianfei Chen, Hang Su et al.ICML 2023 · 136 citations
- NICE: NoIse-modulated Consistency rEgularization for Data-Efficient GANsYao Ni, Piotr KoniuszNeurIPS 2023 · 18 citations
- CHAIN: Enhancing Generalization in Data-Efficient GANs via LipsCHitz Continuity ConstrAIned NormalizationYao Ni, Piotr KoniuszCVPR 2024 · 10 citations
Builds on15
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray et al.ICML 2021 · 6,356 citations
- Video Diffusion ModelsJonathan Ho, Tim Salimans, Alexey A. Gritsenko, William Chan et al.NeurIPS 2022 · 2,948 citations
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine et al.NeurIPS 2020 · 2,345 citations
- Image Generation From Small Datasets via Batch Statistics AdaptationAtsuhiro Noguchi, Tatsuya HaradaICCV 2019 · 211 citations
Related papers
- DigGAN: Discriminator gradIent Gap Regularization for GAN Training with Limited DataTiantian Fang, Ruoyu Sun, Alexander G. SchwingNeurIPS 2022 · 27 citations
- Combating Mode Collapse via Offline Manifold Entropy EstimationHaozhe Liu, Bing Li, Haoqian Wu, Hanbang Liang et al.AAAI 2023 · 18 citations
- Regularizing CNN Transfer Learning With Randomised RegressionYang Zhong, Atsuto MakiCVPR 2020
- Bridging Data Gaps in Diffusion Models with Adversarial Noise-Based Transfer LearningXiyu Wang, Baijiong Lin, Daochang Liu, Ying-Cong Chen et al.ICML 2024 · 8 citations
- On Leveraging Pretrained GANs for Generation with Limited DataMiaoyun Zhao, Yulai Cong, Lawrence CarinICML 2020 · 102 citations
