Where and What? Examining Interpretable Disentangled Representations
Xinqi Zhu, Chang Xu, Dacheng Tao
摘要
Capturing interpretable variations has long been one of the goals in disentanglement learning. However, unlike the independence assumption, interpretability has rarely been exploited to encourage disentanglement in the unsupervised setting. In this paper, we examine the interpretability of disentangled representations by investigating two questions: where to be interpreted and what to be interpreted? A latent code is easily to be interpreted if it would consistently impact a certain subarea of the resulting generated image. We thus propose to learn a spatial mask to localize the effect of each individual latent dimension. On the other hand, interpretability usually comes from latent dimensions that capture simple and basic variations in data. We thus impose a perturbation on a certain dimension of the latent code, and expect to identify the perturbation along this dimension from the generated images so that the encoding of simple variations can be enforced. Additionally, we develop an unsupervised model selection method, which accumulates perceptual distance scores along axes in the latent space. On various datasets, our models can learn high-quality disentangled representations without supervision, showing the proposed modeling of interpretability is an effective proxy for achieving unsupervised disentanglement.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Diffusion Model with Cross Attention as an Inductive Bias for DisentanglementTao Yang, Cuiling Lan, Yan Lu, Nanning ZhengNeurIPS 2024 · 被引用 41 次
- StyleT2I: Toward Compositional and High-Fidelity Text-to-Image SynthesisZhiheng Li, Martin Renqiang Min, Kai Li, Chenliang XuCVPR 2022 · 被引用 38 次
- Commutative Lie Group VAE for Disentanglement LearningXinqi Zhu, Chang Xu, Dacheng TaoICML 2021 · 被引用 35 次
- DyTed: Disentangled Representation Learning for Discrete-time Dynamic GraphKaike Zhang, Qi Cao, Gaolin Fang, Bingbing Xu 等KDD 2023 · 被引用 28 次
- Interactive Disentanglement: Learning Concepts by Interacting with their Prototype RepresentationsWolfgang Stammer, Marius Memmel, Patrick Schramowski, Kristian KerstingCVPR 2022 · 被引用 14 次
它引用的顶会 Paper7
- Weakly-Supervised Disentanglement Without CompromisesFrancesco Locatello, Ben Poole, Gunnar Rätsch, Bernhard Schölkopf 等ICML 2020 · 被引用 361 次
- Theory and Evaluation Metrics for Learning Disentangled RepresentationsKien Do, Truyen TranICLR 2020 · 被引用 107 次
- Unsupervised Model Selection for Variational Disentangled Representation LearningSunny Duan, Loic Matthey, Andre Saraiva, Nick Watters 等ICLR 2020 · 被引用 87 次
- Disentangling Factors of Variations Using Few LabelsFrancesco Locatello, Michael Tschannen, Stefan Bauer, Gunnar Rätsch 等ICLR 2020 · 被引用 77 次
- Progressive Learning and Disentanglement of Hierarchical RepresentationsZhiyuan Li, Jaideep Vitthal Murkute, Prashnna Kumar Gyawali, Linwei WangICLR 2020 · 被引用 47 次
相关 Paper
- Interpretable Generative Adversarial NetworksChao Li, Kelu Yao, Jin Wang, Boyu Diao 等AAAI 2022 · 被引用 19 次
- Learning to Manipulate Individual Objects in an ImageYanchao Yang, Yutong Chen, Stefano SoattoCVPR 2020
- Towards Robust Metrics for Concept Representation EvaluationMateo Espinosa Zarlenga, Pietro Barbiero, Zohreh Shams, Dmitry Kazhdan 等AAAI 2023 · 被引用 32 次
- Factorized Diffusion Autoencoder for Unsupervised Disentangled Representation LearningAncong Wu, Wei-Shi ZhengAAAI 2024 · 被引用 10 次
- InfoGAN-CR and ModelCentrality: Self-supervised Model Training and Selection for Disentangling GANsZinan Lin, Kiran Koshy Thekumparampil, Giulia Fanti, Sewoong OhICML 2020 · 被引用 106 次
