Stacked Homography Transformations for Multi-View Pedestrian Detection
Liangchen Song, Jialian Wu, Ming Yang, Qian Zhang, Yuan Li, Junsong Yuan
摘要
Multi-view pedestrian detection aims to predict a bird’s eye view (BEV) occupancy map from multiple camera views. This task is confronted with two challenges: how to establish the 3D correspondences from views to the BEV map and how to assemble occupancy information across views. In this paper, we propose a novel Stacked HOmography Transformations (SHOT) approach, which is motivated by approximating projections in 3D world coordinates via a stack of homographies. We first construct a stack of transformations for projecting views to the ground plane at different height levels. Then we design a soft selection module so that the network learns to predict the likelihood of the stack of transformations. Moreover, we provide an in-depth theoretical analysis on constructing SHOT and how well SHOT approximates projections in 3D world coordinates. SHOT is empirically verified to be capable of estimating accurate correspondences from individual views to the BEV map, leading to new state-of-the-art performance on standard evaluation benchmarks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- From a Bird's Eye View to See: Joint Camera and Subject Registration without the Camera CalibrationZekun Qian, Ruize Han, Wei Feng, Song WangCVPR 2024 · 被引用 8 次
- Multi-View People Detection in Large Scenes via Supervised View-Wise Contribution WeightingQi Zhang, Yunfei Gong, Daijie Chen, Antoni B. Chan 等AAAI 2024 · 被引用 7 次
- Multi-View Pedestrian Occupancy Prediction with a Novel Synthetic DatasetSithu Aung, Min-Cheol Sagong, Junghyun ChoAAAI 2025 · 被引用 5 次
- Unsupervised Multi-view Pedestrian DetectionMengyin Liu, Chao Zhu, Shiqi Ren, Xu-Cheng YinACM MM 2024 · 被引用 4 次
- MVTrajecter: Multi-View Pedestrian Tracking With Trajectory Motion Cost and Trajectory Appearance CostTaiga Yamane, Ryo Masumura, Satoshi Suzuki, Shota OrihashiICCV 2025 · 被引用 2 次
它引用的顶会 Paper6
- Learnable Triangulation of Human PoseKarim Iskakov, Egor Burkov, Victor S. Lempitsky, Yury MalkovICCV 2019 · 被引用 419 次
- 3D Crowd Counting via Multi-View Fusion with 3D Gaussian KernelsQi Zhang, Antoni B. ChanAAAI 2020 · 被引用 41 次
- Simultaneous Multi-View Instance Detection With Learned Geometric Soft-ConstraintsAhmed Samy Nassar, Sébastien Lefèvre, Jan Dirk WegnerICCV 2019 · 被引用 29 次
- Epipolar TransformersYihui He, Rui Yan, Katerina Fragkiadaki, Shoou-I YuCVPR 2020
- Cross-View Cross-Scene Multi-View Crowd CountingQi Zhang, Wei Lin, Antoni B. ChanCVPR 2021
相关 Paper
- PandaNet: Anchor-Based Single-Shot Multi-Person 3D Pose EstimationAbdallah Benzine, Florian Chabot, Bertrand Luvison, Quoc Cuong Pham 等CVPR 2020
- BEV-SAN: Accurate BEV 3D Object Detection via Slice Attention NetworksXiaowei Chi, Jiaming Liu, Ming Lu, Rongyu Zhang 等CVPR 2023
- Discriminative Spatial Feature Learning for Person Re-IdentificationPeixi Peng, Yonghong Tian, Yangru Huang, Xiangqian Wang 等ACM MM 2020 · 被引用 5 次
- Temporal Enhanced Training of Multi-view 3D Object Detector via Historical Object PredictionZhuofan Zong, Dongzhi Jiang, Guanglu Song, Zeyue Xue 等ICCV 2023 · 被引用 63 次
- OccluBEV: Occlusion Aware Spatiotemporal Modeling for Multi-view 3D Object DetectionZiteng Wen, Hai Xu, Chenyu Liu, Tao Guo 等ACM MM 2023 · 被引用 5 次
