Neural Architecture Search for Joint Human Parsing and Pose Estimation
Dan Zeng, Yuhang Huang, Qian Bao, Junjie Zhang, Chi Su, Wu Liu
摘要
Human parsing and pose estimation are crucial for the understanding of human behaviors. Since these tasks are closely related, employing one unified model to perform two tasks simultaneously allows them to benefit from each other. However, since human parsing is a pixel-wise classification process while pose estimation is usually a regression task, it is non-trivial to extract discriminative features for both tasks while modeling their correlation in the joint learning fashion. Recent studies have shown that Neural Architecture Search (NAS) has the ability to allocate efficient feature connections for specific tasks automatically. With the spirit of NAS, we propose to search for an efficient network architecture (NPPNet) to tackle two tasks at the same time. On the one hand, to extract task-specific features for the two tasks and lay the foundation for the further searching of feature interaction, we propose to search their encoder-decoder architectures, respectively. On the other hand, to ensure two tasks fully communicate with each other, we propose to embed NAS units in both multi-scale feature interaction and high-level feature fusion to establish optimal connections between two tasks. Experimental results on both parsing and pose estimation benchmark datasets have demonstrated that the searched model achieves state-of-the-art performances on both tasks.1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Human Parsing with Joint Learning for Dynamic mmWave Radar Point CloudShuai Wang, Dongjiang Cao, Ruofeng Liu, Wenchao Jiang 等UbiComp 2023 · 被引用 34 次
- REMOT: A Region-to-Whole Framework for Realistic Human Motion TransferQuanwei Yang, Xinchen Liu, Wu Liu, Hongtao Xie 等ACM MM 2022 · 被引用 5 次
- Continuous Heatmap Regression for Pose Estimation via Implicit Neural RepresentationShengxiang Hu, Huaijiang Sun, Dong Wei, Xiaoning Sun 等NeurIPS 2024 · 被引用 5 次
- Semantic Human Parsing via Scalable Semantic Transfer Over Multiple Label DomainsJie Yang, Chaoqun Wang, Zhen Li, Junle Wang 等CVPR 2023
- Fast Adaptation for Human Pose Estimation via Meta-OptimizationShengxiang Hu, Huaijiang Sun, Bin Li, Dong Wei 等CVPR 2024
它引用的顶会 Paper11
- PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture SearchYuhui Xu, Lingxi Xie, Xiaopeng Zhang, Xin Chen 等ICLR 2020 · 被引用 691 次
- Learning Compositional Neural Information Fusion for Human ParsingWenguan Wang, Zhijie Zhang, Siyuan Qi, Jianbing Shen 等ICCV 2019 · 被引用 131 次
- Hybrid Resolution Network Using Edge Guided Region Mutual Information Loss for Human ParsingYunan Liu, Liang Zhao, Shanshan Zhang, Jian YangACM MM 2020 · 被引用 21 次
- Pose-native Network Architecture Search for Multi-person Human Pose EstimationQian Bao, Wu Liu, Jun Hong, Lingyu Duan 等ACM MM 2020 · 被引用 13 次
- Correlating Edge, Pose With ParsingZiwei Zhang, Chi Su, Liang Zheng, Xiaodong XieCVPR 2020
相关 Paper
- ViPNAS: Efficient Video Pose Estimation via Neural Architecture SearchLumin Xu, Yingda Guan, Sheng Jin, Wentao Liu 等CVPR 2021
- HR-NAS: Searching Efficient High-Resolution Neural Architectures With Lightweight TransformersMingyu Ding, Xiaochen Lian, Linjie Yang, Peng Wang 等CVPR 2021
- Differentiable Multi-Granularity Human Representation Learning for Instance-Aware Human Semantic ParsingTianfei Zhou, Wenguan Wang, Si Liu, Yi Yang 等CVPR 2021
- Unified Pose Sequence ModelingLin Geng Foo, Tianjiao Li, Hossein Rahmani, Qiuhong Ke 等CVPR 2023
- Automatic Network Architecture Search for RGB-D Semantic SegmentationWenna Wang, Tao Zhuo, Xiuwei Zhang, Mingjun Sun 等ACM MM 2023 · 被引用 6 次
