HP-Capsule: Unsupervised Face Part Discovery by Hierarchical Parsing Capsule Network
Chang Yu, Xiangyu Zhu, Xiaomei Zhang, Zidu Wang, Zhaoxiang Zhang, Zhen Lei
摘要
Capsule networks are designed to present the objects by a set of parts and their relationships, which provide an insight into the procedure of visual perception. Although recent works have shown the success of capsule networks on simple objects like digits, the human faces with homologous structures, which are suitable for capsules to describe, have not been explored. In this paper, we propose a Hierarchical Parsing Capsule Network (HP-Capsule) for unsupervised face subpart-part discovery. When browsing large-scale face images without labels, the network first encodes the frequently observed patterns with a set of explainable subpart capsules. Then, the subpart capsules are assembled into part-level capsules through a Transformer-based Parsing Module (TPM) to learn the compositional relations between them. During training as the face hierarchy is progressively built and refined, the part capsules adaptively encode the face parts with semantic consistency. HP-Capsule extends the application of capsule networks from digits to human faces and takes a step forward to show how the neural networks understand homologous objects without human intervention. Besides, HP-Capsule gives unsupervised face segmentation results by the covered regions of part capsules, enabling qualitative and quantitative evaluation. Experiments on BP4D and Multi-PIE datasets show the effectiveness of our method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Unsupervised Part Discovery via Descriptor-Based Masked Image Restoration with Optimized ConstraintsJiahao Xia, Yike Wu, Wenjian Huang, Jianguo Zhang 等ICCV 2025 · 被引用 1 次
- Graphics Capsule: Learning Hierarchical 3D Face Representations from 2D ImagesChang Yu, Xiangyu Zhu, Xiaomei Zhang, Zhaoxiang Zhang 等CVPR 2023
- Bridging Neural and Symbolic Representations with Transitional Dictionary LearningJunyan Cheng, Peter ChinICLR 2024
- ParseCaps: An Interpretable Parsing Capsule Network for Medical Image DiagnosisXinyu Geng, Jiaming Wang, Xiaolin Huang, Fanglin Chen 等AAAI 2025
它引用的顶会 Paper17
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Object-Centric Learning with Slot AttentionFrancesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran 等NeurIPS 2020 · 被引用 1,275 次
- Unsupervised Semantic Segmentation by Contrasting Object Mask ProposalsWouter Van Gansbeke, Simon Vandenhende, Stamatios Georgoulis, Luc Van GoolICCV 2021 · 被引用 285 次
- Auto-ReID: Searching for a Part-Aware ConvNet for Person Re-IdentificationRuijie Quan, Xuanyi Dong, Yu Wu, Linchao Zhu 等ICCV 2019 · 被引用 240 次
- Unsupervised Graph Association for Person Re-IdentificationJinlin Wu, Hao Liu, Yang Yang, Zhen Lei 等ICCV 2019 · 被引用 116 次
相关 Paper
- Unsupervised Part Representation by Flow CapsulesSara Sabour, Andrea Tagliasacchi, Soroosh Yazdani, Geoffrey E. Hinton 等ICML 2021 · 被引用 41 次
- PT-CapsNet: A Novel Prediction-Tuning Capsule Network Suitable for Deeper ArchitecturesChenbin Pan, Senem VelipasalarICCV 2021 · 被引用 11 次
- Hierarchical Human Parsing With Typed Part-Relation ReasoningWenguan Wang, Hailong Zhu, Jifeng Dai, Yanwei Pang 等CVPR 2020
- SubSpace Capsule NetworkMarzieh Edraki, Nazanin Rahnavard, Mubarak ShahAAAI 2020 · 被引用 38 次
- Object Part Parsing with Hierarchical Dual TransformerJiamin Chen, Jianlou Si, Naihao Liu, Yao Wu 等ACM MM 2023 · 被引用 1 次
