RSGNet: Relation based Skeleton Graph Network for Crowded Scenes Pose Estimation
Yan Dai, Xuanhan Wang, Lianli Gao, Jingkuan Song, Heng Tao Shen
Abstract
Despite of the recent great progress on multi-person pose estimation, existing solutions still remain challenging under the condition of "crowded scenes", where RGB images capture complex real-world scenes with highly-overlapped people, severe occlusions and diverse postures. In this work, we focus on two main problems: 1) how to design an effective pipeline for crowded scenes pose estimation; and 2) how to equip this pipeline with the ability of relation modeling for interference resolving. To tackle these problems, we propose a new pipeline named Relation based Skeleton Graph Network (RSGNet). Unlike existing works that directly predict joints-of-target by labeling joints-of-interference as false positive, we first encourage all joints to be predicted. And then, a Target-aware Relation Parser (TRP) is designed to model the relation over all predicted joints, resulting in a targetaware encoding. This new pipeline will largely relieve the confusion of the joints estimation model when seeing identical joints with totally distinct labels (e.g., the identical hand exists in two bounding boxes). Furthermore, we introduce a Skeleton Graph Machine (SGM) to model the skeletonbased commonsense knowledge, aiming to estimate the target pose with the constraint of human body structure. Such skeleton-based constraint can help to deal with the challenges in crowded scenes from a reasoning perspective. Solid experiments on pose estimation benchmarks demonstrate that our method outperforms existing state-of-the-art methods. The code and pre-trained models are publicly available online 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1740af50-2a2d-4ee4-9d24-4bb97072fedfCited by top-tier papers1
Ask how each one uses itBuilds on5
- Single-Stage Multi-Person Pose MachinesXuecheng Nie, Jiashi Feng, Jianfeng Zhang, Shuicheng YanICCV 2019 · 246 citations
- KTN: Knowledge Transfer Network for Multi-person DensePose EstimationXuanhan Wang, Lianli Gao, Jingkuan Song, Heng Tao ShenACM MM 2020 · 12 citations
- Distribution-Aware Coordinate Representation for Human Pose EstimationFeng Zhang, Xiatian Zhu, Hanbin Dai, Mao Ye et al.CVPR 2020
- HigherHRNet: Scale-Aware Representation Learning for Bottom-Up Human Pose EstimationBowen Cheng, Bin Xiao, Jingdong Wang, Honghui Shi et al.CVPR 2020
- The Devil Is in the Details: Delving Into Unbiased Data Processing for Human Pose EstimationJunjie Huang, Zheng Zhu, Feng Guo, Guan HuangCVPR 2020
Related papers
- SM-SGE: A Self-Supervised Multi-Scale Skeleton Graph Encoding Framework for Person Re-IdentificationHaocong Rao, Xiping Hu, Jun Cheng, Bin HuACM MM 2021 · 18 citations
- Semantic-aware Transfer with Instance-adaptive Parsing for Crowded Scenes Pose EstimationXuanhan Wang, Lianli Gao, Yan Dai, Yixuan Zhou et al.ACM MM 2021 · 14 citations
- Learning to Estimate Robust 3D Human Mesh from In-the-Wild Crowded ScenesHongsuk Choi, Gyeongsik Moon, JoonKyu Park, Kyoung Mu LeeCVPR 2022 · 92 citations
- Dynamic Graph Reasoning for Multi-person 3D Pose EstimationZhongwei Qiu, Qiansheng Yang, Jian Wang, Dongmei FuACM MM 2022 · 14 citations
- Learning Topology-Aware Dynamic Associations for Robust Multi-Person Pose EstimationShengnan Hu, Yandong Liu, Jiangnan Liu, Yahong ChenAAAI 2026
