Inducing Hierarchical Compositional Model by Sparsifying Generator Network
Xianglei Xing, Tianfu Wu, Song-Chun Zhu, Ying Nian Wu
Abstract
This paper proposes to learn hierarchical compositional AND-OR model for interpretable image synthesis by sparsifying the generator network. The proposed method adopts the scene-objects-parts-subparts-primitives hierarchy in image representation. A scene has different types (i.e., OR) each of which consists of a number of objects (i.e., AND). This can be recursively formulated across the scene-objects-parts-subparts hierarchy and is terminated at the primitive level (e.g., wavelets-like basis). To realize this AND-OR hierarchy in image synthesis, we learn a generator network that consists of the following two components: (i) Each layer of the hierarchy is represented by an overcomplete set of convolutional basis functions. Off-the-shelf convolutional neural architectures are exploited to implement the hierarchy. (ii) Sparsity-inducing constraints are introduced in end-to-end training, which induces a sparsely activated and sparsely connected AND-OR model from the initially densely connected generator network. A straightforward sparsity-inducing constraint is utilized, that is to only allow the top-k basis functions to be activated at each layer (where k is a hyper-parameter). The learned basis functions are also capable of image reconstruction to explain the input images. In experiments, the proposed method is tested on four benchmark datasets. The results show that meaningful and interpretable hierarchical representations are learned with better qualities of image synthesis and reconstruction obtained than baselines.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cc91910e-2ff0-4de0-b13a-6e6b2bfbdf80Cited by top-tier papers2
- Emergence of Shape Bias in Convolutional Neural Networks through Activation SparsityTianqin Li, Ziqi Wen, Yangfan Li, Tai Sing LeeNeurIPS 2023 · 24 citations
- HiABP: Hierarchical Initialized ABP for Unsupervised Representation LearningJiankai Sun, Rui Liu, Bolei ZhouAAAI 2021 · 3 citations
Related papers
- Hierarchical Concept Embedding & Pursuit for Interpretable Image ClassificationNghia Nguyen, Tianjiao Ding, Rene VidalCVPR 2026 · 1 citation
- Generative Scene Graph NetworksFei Deng, Zhuo Zhi, Donghun Lee, Sungjin AhnICLR 2021 · 10 citations
- Towards Interpretable Object Detection by Unfolding Latent StructuresTianfu Wu, Xi SongICCV 2019 · 28 citations
- Learning by Analogy: A Causal Framework for Compositional GeneralizationLingjing Kong, Shaoan Xie, Yang Jiao, Yetian Chen et al.CVPR 2026
- Learning Unsupervised Hierarchical Part Decomposition of 3D Objects From a Single RGB ImageDespoina Paschalidou, Luc Van Gool, Andreas GeigerCVPR 2020
