Decoupling Global and Local Representations via Invertible Generative Flows
Xuezhe Ma, Xiang Kong, Shanghang Zhang, Eduard H. Hovy
Abstract
In this work, we propose a new generative model that is capable of automatically decoupling global and local representations of images in an entirely unsupervised setting, by embedding a generative flow in the VAE framework to model the decoder. Specifically, the proposed model utilizes the variational auto-encoding framework to learn a (low-dimensional) vector of latent variables to capture the global information of an image, which is fed as a conditional input to a flow-based invertible decoder with architecture borrowed from style transfer literature. Experimental results on standard image benchmarks demonstrate the effectiveness of our model in terms of density estimation, image generation and unsupervised representation learning. Importantly, this work demonstrates that with only architectural inductive biases, a generative model with a likelihood-based objective is capable of learning decoupled representations, requiring no explicit supervision. The code for our model is available at https://github.com/XuezheMax/wolf .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 965d67c9-2ae7-4dd6-a9db-cde322b5bff0Cited by top-tier papers5
- Maximum Likelihood Training of Implicit Nonlinear Diffusion ModelDongjun Kim, Byeonghu Na, Se Jung Kwon, Dongsoo Lee et al.NeurIPS 2022 · 61 citations
- Better May Not Be Fairer: A Study on Subgroup Discrepancy in Image ClassificationMing-Chang Chiu, Pin-Yu Chen, Xuezhe MaICCV 2023 · 9 citations
- CHIMLE: Conditional Hierarchical IMLE for Multimodal Conditional Image SynthesisShichong Peng, Seyed Alireza Moazenipourasil, Ke LiNeurIPS 2022 · 4 citations
- DifAttack: Query-Efficient Black-Box Adversarial Attack via Disentangled Feature SpaceJun Liu, Jiantao Zhou, Jiandian Zeng, Jinyu TianAAAI 2024 · 2 citations
- Improving Unsupervised Hierarchical Representation With Reinforcement LearningRuyi An, Yewen Li, Xu He, Pengjie Gu et al.CVPR 2024
Builds on3
- Relaxing Bijectivity Constraints with Continuously Indexed Normalising FlowsRobert Cornish, Anthony L. Caterini, George Deligiannidis, Arnaud DoucetICML 2020 · 141 citations
- VFlow: More Expressive Generative Flows with Variational Data AugmentationJianfei Chen, Cheng Lu, Biqi Chenli, Jun Zhu et al.ICML 2020 · 64 citations
- Latent Normalizing Flows for Many-to-Many Cross-Domain MappingsShweta Mahajan, Iryna Gurevych, Stefan RothICLR 2020 · 38 citations
Related papers
- Structure by Architecture: Structured Representations without RegularizationFelix Leeb, Giulia Lanzillotta, Yashas Annadani, Michel Besserve et al.ICLR 2023 · 1 citation
- Autoregressive Stylized Motion Synthesis With Generative FlowYu-Hui Wen, Zhipeng Yang, Hongbo Fu, Lin Gao et al.CVPR 2021
- ShapeFlow: Learnable Deformation Flows Among 3D ShapesChiyu Max Jiang, Jingwei Huang, Andrea Tagliasacchi, Leonidas J. GuibasNeurIPS 2020 · 46 citations
- Retriever: Learning Content-Style Representation as a Token-Level Bipartite GraphDacheng Yin, Xuanchi Ren, Chong Luo, Yuwang Wang et al.ICLR 2022 · 13 citations
- Adversarial Disentanglement with Grouped ObservationsJózsef NémethAAAI 2020 · 8 citations
