Cross-View Tracking for Multi-Human 3D Pose Estimation at Over 100 FPS
Long Chen, Haizhou Ai, Rui Chen, Zijie Zhuang, Shuang Liu
摘要
Estimating 3D poses of multiple humans in real-time is a classic but still challenging task in computer vision. Its major difficulty lies in the ambiguity in cross-view association of 2D poses and the huge state space when there are multiple people in multiple views. In this paper, we present a novel solution for multi-human 3D pose estimation from multiple calibrated camera views. It takes 2D poses in different camera coordinates as inputs and aims for the accurate 3D poses in the global coordinate. Unlike previous methods that associate 2D poses among all pairs of views from scratch at every frame, we exploit the temporal consistency in videos to match the 2D inputs with 3D poses directly in 3-space. More specifically, we propose to retain the 3D pose for each person and update them iteratively via the cross-view multi-human tracking. This novel formulation improves both accuracy and efficiency, as we demonstrated on widely-used public datasets. To further verify the scalability of our method, we propose a new large-scale multi-human dataset with 12 to 28 camera views. Without bells and whistles, our solution achieves 154 FPS on 12 cameras and 34 FPS on 28 cameras, indicating its ability to handle large-scale real-world applications. The proposed dataset is released at https://github. com/longcw/crossview_3d_pose_tracking .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- TransPose: real-time 3D human translation and pose estimation with six inertial sensorsXinyu Yi, Yuxiao Zhou, Feng XuSIGGRAPH 2021 · 被引用 200 次
- Physical Inertial Poser (PIP): Physics-aware Real-time Human Motion Tracking from Sparse Inertial SensorsXinyu Yi, Yuxiao Zhou, Marc Habermann, Soshi Shimada 等CVPR 2022 · 被引用 198 次
- EgoLocate: Real-time Motion Capture, Localization, and Mapping with Sparse Body-mounted SensorsXinyu Yi, Yuxiao Zhou, Marc Habermann, Vladislav Golyanik 等SIGGRAPH 2023 · 被引用 62 次
- EgoHumans: An Egocentric 3D Multi-Human BenchmarkRawal Khirodkar, Aayush Bansal, Lingni Ma, Richard A. Newcombe 等ICCV 2023 · 被引用 59 次
- MetaPose: Fast 3D Pose from Multiple Views without 3D SupervisionBen Usman, Andrea Tagliasacchi, Kate Saenko, Avneesh SudCVPR 2022 · 被引用 29 次
它引用的顶会 Paper2
相关 Paper
- Shape-aware Multi-Person Pose Estimation from Multi-View ImagesZijian Dong, Jie Song, Xu Chen, Chen Guo 等ICCV 2021 · 被引用 47 次
- 3D Human Pose Estimation from Multiple Dynamic Views via Single-view Pretraining with Procrustes AlignmentRenshu Gu, Jiajun Zhu, Yixuan Si, Fei Gao 等ACM MM 2024 · 被引用 1 次
- CoMotion: Concurrent Multi-person 3D MotionAlejandro Newell, Peiyun Hu, Lahav Lipson, Stephan R. Richter 等ICLR 2025
- QuickPose: Real-time Multi-view Multi-person Pose Estimation in Crowded ScenesZhize Zhou, Qing Shuai, Yize Wang, Qi Fang 等SIGGRAPH 2022 · 被引用 15 次
- Multi-View Multi-Person 3D Pose Estimation With Plane Sweep StereoJiahao Lin, Gim Hee LeeCVPR 2021
