Learning Compositional Neural Information Fusion for Human Parsing
Wenguan Wang, Zhijie Zhang, Siyuan Qi, Jianbing Shen, Yanwei Pang, Ling Shao
Abstract
This work proposes to combine neural networks with the compositional hierarchy of human bodies for efficient and complete human parsing. We formulate the approach as a neural information fusion framework. Our model assembles the information from three inference processes over the hierarchy: direct inference (directly predicting each part of a human body using image information), bottom-up inference (assembling knowledge from constituent parts), and top-down inference (leveraging context from parent nodes). The bottom-up and top-down inferences explicitly model the compositional and decompositional relations in human bodies, respectively. In addition, the fusion of multi-source information is conditioned on the inputs, i.e., by estimating and considering the confidence of the sources. The whole model is end-to-end differentiable, explicitly modeling information flows and structures. Our approach is extensively evaluated on four popular datasets, outperforming the state-of-the-arts in all cases, with a fast processing speed of 23fps. Our code and results have been released to help ease future research in this direction.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9bc74bf8-3051-4243-b4cb-71c3f80b40adCited by top-tier papers25
- GMMSeg: Gaussian Mixture based Generative Semantic Segmentation ModelsChen Liang, Wenguan Wang, Jiaxu Miao, Yi YangNeurIPS 2022 · 185 citations
- Deep Hierarchical Semantic SegmentationLiulei Li, Tianfei Zhou, Wenguan Wang, Jianwu Li et al.CVPR 2022 · 181 citations
- Group-Wise Semantic Mining for Weakly Supervised Semantic SegmentationXueyi Li, Tianfei Zhou, Jianwu Li, Yi Zhou et al.AAAI 2021 · 143 citations
- PLIP: Language-Image Pre-training for Person Representation LearningJialong Zuo, Jiahao Hong, Feng Zhang, Changqian Yu et al.NeurIPS 2024 · 96 citations
- Logic-induced Diagnostic Reasoning for Semi-supervised Semantic SegmentationChen Liang, Wenguan Wang, Jiaxu Miao, Yi YangICCV 2023 · 55 citations
Related papers
- Hierarchical Human Parsing With Typed Part-Relation ReasoningWenguan Wang, Hailong Zhu, Jifeng Dai, Yanwei Pang et al.CVPR 2020
- Differentiable Multi-Granularity Human Representation Learning for Instance-Aware Human Semantic ParsingTianfei Zhou, Wenguan Wang, Si Liu, Yi Yang et al.CVPR 2021
- Hybrid Resolution Network Using Edge Guided Region Mutual Information Loss for Human ParsingYunan Liu, Liang Zhao, Shanshan Zhang, Jian YangACM MM 2020 · 21 citations
- Complete 3D Human Reconstruction from a Single Incomplete ImageJunying Wang, Jae Shin Yoon, Tuanfeng Y. Wang, Krishna Kumar Singh et al.CVPR 2023
- CDGNet: Class Distribution Guided Network for Human ParsingKunliang Liu, Ouk Choi, Jianming Wang, Wonjun HwangCVPR 2022 · 44 citations
