HP-Capsule: Unsupervised Face Part Discovery by Hierarchical Parsing Capsule Network
Chang Yu, Xiangyu Zhu, Xiaomei Zhang, Zidu Wang, Zhaoxiang Zhang, Zhen Lei
Abstract
Capsule networks are designed to present the objects by a set of parts and their relationships, which provide an insight into the procedure of visual perception. Although recent works have shown the success of capsule networks on simple objects like digits, the human faces with homologous structures, which are suitable for capsules to describe, have not been explored. In this paper, we propose a Hierarchical Parsing Capsule Network (HP-Capsule) for unsupervised face subpart-part discovery. When browsing large-scale face images without labels, the network first encodes the frequently observed patterns with a set of explainable subpart capsules. Then, the subpart capsules are assembled into part-level capsules through a Transformer-based Parsing Module (TPM) to learn the compositional relations between them. During training as the face hierarchy is progressively built and refined, the part capsules adaptively encode the face parts with semantic consistency. HP-Capsule extends the application of capsule networks from digits to human faces and takes a step forward to show how the neural networks understand homologous objects without human intervention. Besides, HP-Capsule gives unsupervised face segmentation results by the covered regions of part capsules, enabling qualitative and quantitative evaluation. Experiments on BP4D and Multi-PIE datasets show the effectiveness of our method.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f23a5347-c78b-4668-ac80-4c4989a98be2Cited by top-tier papers4
- Unsupervised Part Discovery via Descriptor-Based Masked Image Restoration with Optimized ConstraintsJiahao Xia, Yike Wu, Wenjian Huang, Jianguo Zhang et al.ICCV 2025 · 1 citation
- Graphics Capsule: Learning Hierarchical 3D Face Representations from 2D ImagesChang Yu, Xiangyu Zhu, Xiaomei Zhang, Zhaoxiang Zhang et al.CVPR 2023
- Bridging Neural and Symbolic Representations with Transitional Dictionary LearningJunyan Cheng, Peter ChinICLR 2024
- ParseCaps: An Interpretable Parsing Capsule Network for Medical Image DiagnosisXinyu Geng, Jiaming Wang, Xiaolin Huang, Fanglin Chen et al.AAAI 2025
Builds on17
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Object-Centric Learning with Slot AttentionFrancesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran et al.NeurIPS 2020 · 1,275 citations
- Unsupervised Semantic Segmentation by Contrasting Object Mask ProposalsWouter Van Gansbeke, Simon Vandenhende, Stamatios Georgoulis, Luc Van GoolICCV 2021 · 285 citations
- Auto-ReID: Searching for a Part-Aware ConvNet for Person Re-IdentificationRuijie Quan, Xuanyi Dong, Yu Wu, Linchao Zhu et al.ICCV 2019 · 240 citations
- Unsupervised Graph Association for Person Re-IdentificationJinlin Wu, Hao Liu, Yang Yang, Zhen Lei et al.ICCV 2019 · 116 citations
Related papers
- Unsupervised Part Representation by Flow CapsulesSara Sabour, Andrea Tagliasacchi, Soroosh Yazdani, Geoffrey E. Hinton et al.ICML 2021 · 41 citations
- PT-CapsNet: A Novel Prediction-Tuning Capsule Network Suitable for Deeper ArchitecturesChenbin Pan, Senem VelipasalarICCV 2021 · 11 citations
- Hierarchical Human Parsing With Typed Part-Relation ReasoningWenguan Wang, Hailong Zhu, Jifeng Dai, Yanwei Pang et al.CVPR 2020
- SubSpace Capsule NetworkMarzieh Edraki, Nazanin Rahnavard, Mubarak ShahAAAI 2020 · 38 citations
- Object Part Parsing with Hierarchical Dual TransformerJiamin Chen, Jianlou Si, Naihao Liu, Yao Wu et al.ACM MM 2023 · 1 citation
