Differentiable Multi-Granularity Human Representation Learning for Instance-Aware Human Semantic Parsing
Tianfei Zhou, Wenguan Wang, Si Liu, Yi Yang, Luc Van Gool
Abstract
To address the challenging task of instance-aware human part parsing, a new bottom-up regime is proposed to learn category-level human semantic segmentation as well as multi-person pose estimation in a joint and end-to-end manner. It is a compact, efficient and powerful framework that exploits structural information over different human granularities and eases the difficulty of person partitioning. Specifically, a dense-to-sparse projection field, which allows explicitly associating dense human semantics with sparse keypoints, is learnt and progressively improved over the network feature pyramid for robustness. Then, the difficult pixel grouping problem is cast as an easier, multiperson joint assembling task. By formulating joint association as maximum-weight bipartite matching, a differentiable solution is developed to exploit projected gradient descent and Dykstra's cyclic projection algorithm. This makes our method end-to-end trainable and allows back-propagating the grouping error to directly supervise multi-granularity human representation learning. This is distinguished from current bottom-up human parsers or pose estimators which require sophisticated post-processing or heuristic greedy algorithms. Experiments on three instance-aware human parsing datasets show that our model outperforms other bottom-up alternatives with much more efficient inference.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers10
- Regional Semantic Contrast and Aggregation for Weakly Supervised Semantic SegmentationTianfei Zhou, Meijie Zhang, Fang Zhao, Jianwu LiCVPR 2022 · 190 citations
- Deep Hierarchical Semantic SegmentationLiulei Li, Tianfei Zhou, Wenguan Wang, Jianwu Li et al.CVPR 2022 · 181 citations
- Learning Equivariant Segmentation with Instance-Unique QueryingWenguan Wang, James Liang, Dongfang LiuNeurIPS 2022 · 99 citations
- CLUSTSEG: Clustering for Universal SegmentationJames Chenhao Liang, Tianfei Zhou, Dongfang Liu, Wenguan WangICML 2023 · 85 citations
- Going Denser with Open-Vocabulary Part SegmentationPeize Sun, Shoufa Chen, Chenchen Zhu, Fanyi Xiao et al.ICCV 2023 · 83 citations
Builds on12
- RepPoints: Point Set Representation for Object DetectionZe Yang, Shaohui Liu, Han Hu, Liwei Wang et al.ICCV 2019 · 1,056 citations
- CARAFE: Content-Aware ReAssembly of FEaturesJiaqi Wang, Kai Chen, Rui Xu, Ziwei Liu et al.ICCV 2019 · 842 citations
- Human-Aware Motion DeblurringZiyi Shen, Wenguan Wang, Xiankai Lu, Jianbing Shen et al.ICCV 2019 · 374 citations
- TensorMask: A Foundation for Dense Object SegmentationXinlei Chen, Ross B. Girshick, Kaiming He, Piotr DollárICCV 2019 · 357 citations
- Single-Stage Multi-Person Pose MachinesXuecheng Nie, Jiashi Feng, Jianfeng Zhang, Shuicheng YanICCV 2019 · 246 citations
Related papers
- Semantic-aware Transfer with Instance-adaptive Parsing for Crowded Scenes Pose EstimationXuanhan Wang, Lianli Gao, Yan Dai, Yixuan Zhou et al.ACM MM 2021 · 14 citations
- Single-Stage Multi-human Parsing via Point Sets and Center-based OffsetsJiaming Chu, Lei Jin, Xiaojin Fan, Yinglei Teng et al.ACM MM 2023 · 14 citations
- QueryPose: Sparse Multi-Person Pose Regression via Spatial-Aware Part-Level QueryYabo Xiao, Kai Su, Xiaojuan Wang, Dongdong Yu et al.NeurIPS 2022 · 32 citations
- Learning Compositional Neural Information Fusion for Human ParsingWenguan Wang, Zhijie Zhang, Siyuan Qi, Jianbing Shen et al.ICCV 2019 · 131 citations
- Grapy-ML: Graph Pyramid Mutual Learning for Cross-Dataset Human ParsingHaoyu He, Jing Zhang, Qiming Zhang, Dacheng TaoAAAI 2020 · 65 citations
