ComGAN: Unsupervised Disentanglement and Segmentation via Image Composition
Rui Ding, Kehua Guo, Xiangyuan Zhu, Zheng Wu, Liwei Wang
Abstract
We propose ComGAN, a simple unsupervised generative model, which simultaneously generates realistic images and high semantic masks under an adversarial loss and a binary regularization. In this paper, we first investigate two kinds of trivial solutions in the compositional generation process, and demonstrate their source is vanishing gradients on the mask. Then, we solve trivial solutions from the perspective of architecture. Furthermore, we redesign two fully unsupervised modules based on ComGAN (DS-ComGAN), where the disentanglement module associates the foreground, background and mask with three independent variables, and the segmentation module learns object segmentation. Experimental results show that (i) ComGAN's network architecture effectively avoids trivial solutions without any supervised information and regularization; (ii) DS-ComGAN achieves remarkable results and outperforms existing semi-supervised and weakly supervised methods by a large margin in both the image disentanglement and unsupervised segmentation tasks. It implies that the redesign of ComGAN is a possible direction for future unsupervised work. 1 * Corresponding author 1 Code and data are available at https://github.com/Ruiding1/ComGAN 36th Conference on Neural Information Processing Systems (NeurIPS 2022).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on19
- Object-Centric Learning with Slot AttentionFrancesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran et al.NeurIPS 2020 · 1,275 citations
- Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?Rameen Abdal, Yipeng Qin, Peter WonkaICCV 2019 · 1,195 citations
- Unsupervised Discovery of Interpretable Directions in the GAN Latent SpaceAndrey Voynov, Artem BabenkoICML 2020 · 459 citations
- Counterfactual Generative NetworksAxel Sauer, Andreas GeigerICLR 2021 · 145 citations
- InfoGAN-CR and ModelCentrality: Self-supervised Model Training and Selection for Disentangling GANsZinan Lin, Kiran Koshy Thekumparampil, Giulia Fanti, Sewoong OhICML 2020 · 106 citations
Related papers
- Labels4Free: Unsupervised Segmentation using StyleGANRameen Abdal, Peihao Zhu, Niloy J. Mitra, Peter WonkaICCV 2021 · 88 citations
- Semi-Supervised Single-Stage Controllable GANs for Conditional Fine-Grained Image GenerationTianyi Chen, Yi Liu, Yunfei Zhang, Si Wu et al.ICCV 2021 · 11 citations
- Factorized Diffusion Architectures for Unsupervised Image Generation and SegmentationXin Yuan, Michael MaireNeurIPS 2024 · 4 citations
- CoMoGAN: Continuous Model-Guided Image-to-Image TranslationFabio Pizzati, Pietro Cerri, Raoul de CharetteCVPR 2021
- Mask-Embedded Discriminator With Region-Based Semantic Regularization for Semi-Supervised Class-Conditional Image SynthesisYi Liu, Xiaoyang Huo, Tianyi Chen, Xiangping Zeng et al.CVPR 2021
