Where and What? Examining Interpretable Disentangled Representations
Xinqi Zhu, Chang Xu, Dacheng Tao
Abstract
Capturing interpretable variations has long been one of the goals in disentanglement learning. However, unlike the independence assumption, interpretability has rarely been exploited to encourage disentanglement in the unsupervised setting. In this paper, we examine the interpretability of disentangled representations by investigating two questions: where to be interpreted and what to be interpreted? A latent code is easily to be interpreted if it would consistently impact a certain subarea of the resulting generated image. We thus propose to learn a spatial mask to localize the effect of each individual latent dimension. On the other hand, interpretability usually comes from latent dimensions that capture simple and basic variations in data. We thus impose a perturbation on a certain dimension of the latent code, and expect to identify the perturbation along this dimension from the generated images so that the encoding of simple variations can be enforced. Additionally, we develop an unsupervised model selection method, which accumulates perceptual distance scores along axes in the latent space. On various datasets, our models can learn high-quality disentangled representations without supervision, showing the proposed modeling of interpretability is an effective proxy for achieving unsupervised disentanglement.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers12
- Diffusion Model with Cross Attention as an Inductive Bias for DisentanglementTao Yang, Cuiling Lan, Yan Lu, Nanning ZhengNeurIPS 2024 · 41 citations
- StyleT2I: Toward Compositional and High-Fidelity Text-to-Image SynthesisZhiheng Li, Martin Renqiang Min, Kai Li, Chenliang XuCVPR 2022 · 38 citations
- Commutative Lie Group VAE for Disentanglement LearningXinqi Zhu, Chang Xu, Dacheng TaoICML 2021 · 35 citations
- DyTed: Disentangled Representation Learning for Discrete-time Dynamic GraphKaike Zhang, Qi Cao, Gaolin Fang, Bingbing Xu et al.KDD 2023 · 28 citations
- Interactive Disentanglement: Learning Concepts by Interacting with their Prototype RepresentationsWolfgang Stammer, Marius Memmel, Patrick Schramowski, Kristian KerstingCVPR 2022 · 14 citations
Builds on7
- Weakly-Supervised Disentanglement Without CompromisesFrancesco Locatello, Ben Poole, Gunnar Rätsch, Bernhard Schölkopf et al.ICML 2020 · 361 citations
- Theory and Evaluation Metrics for Learning Disentangled RepresentationsKien Do, Truyen TranICLR 2020 · 107 citations
- Unsupervised Model Selection for Variational Disentangled Representation LearningSunny Duan, Loic Matthey, Andre Saraiva, Nick Watters et al.ICLR 2020 · 87 citations
- Disentangling Factors of Variations Using Few LabelsFrancesco Locatello, Michael Tschannen, Stefan Bauer, Gunnar Rätsch et al.ICLR 2020 · 77 citations
- Progressive Learning and Disentanglement of Hierarchical RepresentationsZhiyuan Li, Jaideep Vitthal Murkute, Prashnna Kumar Gyawali, Linwei WangICLR 2020 · 47 citations
Related papers
- Interpretable Generative Adversarial NetworksChao Li, Kelu Yao, Jin Wang, Boyu Diao et al.AAAI 2022 · 19 citations
- Learning to Manipulate Individual Objects in an ImageYanchao Yang, Yutong Chen, Stefano SoattoCVPR 2020
- Towards Robust Metrics for Concept Representation EvaluationMateo Espinosa Zarlenga, Pietro Barbiero, Zohreh Shams, Dmitry Kazhdan et al.AAAI 2023 · 32 citations
- Factorized Diffusion Autoencoder for Unsupervised Disentangled Representation LearningAncong Wu, Wei-Shi ZhengAAAI 2024 · 10 citations
- InfoGAN-CR and ModelCentrality: Self-supervised Model Training and Selection for Disentangling GANsZinan Lin, Kiran Koshy Thekumparampil, Giulia Fanti, Sewoong OhICML 2020 · 106 citations
