Compressed Volumetric Heatmaps for Multi-Person 3D Pose Estimation
Matteo Fabbri, Fabio Lanzi, Simone Calderara, Stefano Alletto, Rita Cucchiara
摘要
In this paper we present a novel approach for bottomup multi-person 3D human pose estimation from monocular RGB images. We propose to use high resolution volumetric heatmaps to model joint locations, devising a simple and effective compression method to drastically reduce the size of this representation. At the core of the proposed method lies our Volumetric Heatmap Autoencoder, a fully-convolutional network tasked with the compression of ground-truth heatmaps into a dense intermediate representation. A second model, the Code Predictor, is then trained to predict these codes, which can be decompressed at test time to re-obtain the original representation. Our experimental evaluation shows that our method performs favorably when compared to state of the art on both multi-person and single-person 3D human pose estimation datasets and, thanks to our novel compression strategy, can process full-HD images at the constant runtime of 8 fps regardless of the number of subjects in the scene. Code and models available at https://github.com/fabbrimatteo/LoCO . + × e-c2d e-c3d d -c3d d -c2d feature extractor f-c2d L2 Encoder e d Decoder Code Predictor f Code Predictor (train) f e VHA (train and eval.) d e f D'×H ''×W '' N×D '×H ''×W '' N×D '×H ''×W '' D '×H ''×W '' D '×H ''×W '' 3×H×W N×D×H '×W ' N×D×H '×W ' + concat. × deconcat. Code Predictor (eval.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper19
- MOTSynth: How Can Synthetic Data Help Pedestrian Detection and Tracking?Matteo Fabbri, Guillem Brasó, Gianluca Maugeri, Orcun Cetintas 等ICCV 2021 · 被引用 128 次
- GLAMR: Global Occlusion-Aware Human Mesh Recovery with Dynamic CamerasYe Yuan, Umar Iqbal, Pavlo Molchanov, Kris Kitani 等CVPR 2022 · 被引用 111 次
- Neural monocular 3D human motion capture with physical awarenessSoshi Shimada, Vladislav Golyanik, Weipeng Xu, Patrick Pérez 等SIGGRAPH 2021 · 被引用 107 次
- Improving Robustness and Accuracy via Relative Information Encoding in 3D Human Pose EstimationWenkang Shan, Haopeng Lu, Shanshe Wang, Xinfeng Zhang 等ACM MM 2021 · 被引用 62 次
- Shape-aware Multi-Person Pose Estimation from Multi-View ImagesZijian Dong, Jie Song, Xu Chen, Chen Guo 等ICCV 2021 · 被引用 47 次
它引用的顶会 Paper1
相关 Paper
- Distribution-Aware Single-Stage Models for Multi-Person 3D Pose EstimationZitian Wang, Xuecheng Nie, Xiaochao Qu, Yunpeng Chen 等CVPR 2022 · 被引用 44 次
- TEMPO: Efficient Multi-View Pose Estimation, Tracking, and ForecastingRohan Choudhury, Kris M. Kitani, László A. JeniICCV 2023 · 被引用 32 次
- PandaNet: Anchor-Based Single-Shot Multi-Person 3D Pose EstimationAbdallah Benzine, Florian Chabot, Bertrand Luvison, Quoc Cuong Pham 等CVPR 2020
- SelfPose3d: Self-Supervised Multi-Person Multi-View 3d Pose EstimationVinkle Srivastav, Keqi Chen, Nicolas PadoyCVPR 2024 · 被引用 17 次
- Mutual Adaptive Reasoning for Monocular 3D Multi-Person Pose EstimationJuze Zhang, Jingya Wang, Ye Shi, Fei Gao 等ACM MM 2022 · 被引用 15 次
