Teachers Do More Than Teach: Compressing Image-to-Image Models
Qing Jin, Jian Ren, Oliver J. Woodford, Jiazhuo Wang, Geng Yuan, Yanzhi Wang, Sergey Tulyakov
摘要
Generative Adversarial Networks (GANs) have achieved huge success in generating high-fidelity images, however, they suffer from low efficiency due to tremendous computational cost and bulky memory usage. Recent efforts on compression GANs show noticeable progress in obtaining smaller generators by sacrificing image quality or involving a time-consuming searching process. In this work, we aim to address these issues by introducing a teacher network that provides a search space in which efficient network architectures can be found, in addition to performing knowledge distillation. First, we revisit the search space of generative models, introducing an inception-based residual block into generators. Second, to achieve target computation cost, we propose a one-step pruning algorithm that searches a student architecture from the teacher model and substantially reduces searching cost. It requires no ℓ 1 sparsity regularization and its associated hyper-parameters, simplifying the training procedure. Finally, we propose to distill knowledge through maximizing feature similarity between teacher and student via an index named Global Kernel Alignment (GKA). Our compressed networks achieve similar or even better image fidelity (FID, mIoU) than the original models with much-reduced computational cost, e.g., MACs. Code will be released at https://github.com/snap-research/CAT .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- Rethinking Vision Transformers for MobileNet Size and SpeedYanyu Li, Ju Hu, Yang Wen, Georgios Evangelidis 等ICCV 2023 · 被引用 300 次
- SnapFusion: Text-to-Image Diffusion Model on Mobile Devices within Two SecondsYanyu Li, Huan Wang, Qing Jin, Ju Hu 等NeurIPS 2023 · 被引用 300 次
- A Good Image Generator Is What You Need for High-Resolution Video SynthesisYu Tian, Jian Ren, Menglei Chai, Kyle Olszewski 等ICLR 2021 · 被引用 208 次
- Wavelet Knowledge Distillation: Towards Efficient Image-to-Image TranslationLinfeng Zhang, Xin Chen, Xiaobing Tu, Pengfei Wan 等CVPR 2022 · 被引用 105 次
- MobileFaceSwap: A Lightweight Framework for Video Face SwappingZhiliang Xu, Zhibin Hong, Changxing Ding, Zhen Zhu 等AAAI 2022 · 被引用 78 次
它引用的顶会 Paper17
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang 等ICLR 2020 · 被引用 1,522 次
- Progressive Differentiable Architecture Search: Bridging the Depth Gap Between Search and EvaluationXin Chen, Lingxi Xie, Jun Wu, Qi TianICCV 2019 · 被引用 725 次
- Universally Slimmable Networks and Improved Training TechniquesJiahui Yu, Thomas S. HuangICCV 2019 · 被引用 444 次
- PatDNN: Achieving Real-Time DNN Execution on Mobile Devices with Pattern-based Weight PruningWei Niu, Xiaolong Ma, Sheng Lin, Shihao Wang 等ASPLOS 2020 · 被引用 214 次
- A Good Image Generator Is What You Need for High-Resolution Video SynthesisYu Tian, Jian Ren, Menglei Chai, Kyle Olszewski 等ICLR 2021 · 被引用 208 次
相关 Paper
- Discriminator-Cooperated Feature Map Distillation for GAN CompressionTie Hu, Mingbao Lin, Lizhou You, Fei Chao 等CVPR 2023
- Information-Theoretic GAN Compression with Variational Energy-based ModelMinsoo Kang, Hyewon Yoo, Eunhee Kang, Sehwan Ki 等NeurIPS 2022 · 被引用 5 次
- Distilling Portable Generative Adversarial Networks for Image TranslationHanting Chen, Yunhe Wang, Han Shu, Changyuan Wen 等AAAI 2020 · 被引用 89 次
- Teacher Guided Neural Architecture Search for Face RecognitionXiaobo WangAAAI 2021 · 被引用 12 次
- Content-Aware GAN CompressionYuchen Liu, Zhixin Shu, Yijun Li, Zhe Lin 等CVPR 2021
