Information-Theoretic GAN Compression with Variational Energy-based Model
Minsoo Kang, Hyewon Yoo, Eunhee Kang, Sehwan Ki, Hyong-Euk Lee, Bohyung Han
Abstract
We propose an information-theoretic knowledge distillation approach for the compression of generative adversarial networks, which aims to maximize the mutual information between teacher and student networks via a variational optimization based on an energy-based model. Because the direct computation of the mutual information in continuous domains is intractable, our approach alternatively optimizes the student network by maximizing the variational lower bound of the mutual information. To achieve a tight lower bound, we introduce an energy-based model relying on a deep neural network to represent a flexible variational distribution that deals with high-dimensional images and consider spatial dependencies between pixels, effectively. Since the proposed method is a generic optimization algorithm, it can be conveniently incorporated into arbitrary generative adversarial networks and even dense prediction networks, e.g., image enhancement models. We demonstrate that the proposed algorithm achieves outstanding performance in model compression of generative adversarial networks consistently when combined with several existing models. Existing approaches for compressing GANs [31, 35, 32, [36] [37] [38] 33, 39] often utilize knowledge distillation and manage to achieve competitive accuracy of student models. Specifically, these methods transfer the knowledge of a teacher network to a student model via intermediate representations, 36th Conference on Neural Information Processing Systems (NeurIPS 2022).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ba480ca1-d0cb-42be-bfa2-d9217fdb01b3Cited by top-tier papers1
Ask how each one uses itBuilds on15
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang et al.ICLR 2020 · 1,522 citations
- Contrastive Representation DistillationYonglong Tian, Dilip Krishnan, Phillip IsolaICLR 2020 · 1,305 citations
- Operation-Aware Soft Channel Pruning using Differentiable MasksMinsoo Kang, Bohyung HanICML 2020 · 165 citations
- COT-GAN: Generating Sequential Data via Causal Optimal TransportTianlin Xu, Li Kevin Wenliang, Michael Munn, Beatrice AcciaioNeurIPS 2020 · 139 citations
- Co-Evolutionary Compression for Unpaired Image TranslationHan Shu, Yunhe Wang, Xu Jia, Kai Han et al.ICCV 2019 · 93 citations
Related papers
- Distilling Portable Generative Adversarial Networks for Image TranslationHanting Chen, Yunhe Wang, Han Shu, Changyuan Wen et al.AAAI 2020 · 89 citations
- AutoGAN-Distiller: Searching to Compress Generative Adversarial NetworksYonggan Fu, Wuyang Chen, Haotao Wang, Haoran Li et al.ICML 2020 · 91 citations
- Self-Supervised Generative Adversarial CompressionChong Yu, Jeff PoolNeurIPS 2020 · 15 citations
- Teachers Do More Than Teach: Compressing Image-to-Image ModelsQing Jin, Jian Ren, Oliver J. Woodford, Jiazhuo Wang et al.CVPR 2021
- Discriminator-Cooperated Feature Map Distillation for GAN CompressionTie Hu, Mingbao Lin, Lizhou You, Fei Chao et al.CVPR 2023
