Simultaneous Multi-View Instance Detection With Learned Geometric Soft-Constraints
Ahmed Samy Nassar, Sébastien Lefèvre, Jan Dirk Wegner
摘要
We propose to jointly learn multi-view geometry and warping between views of the same object instances for robust cross-view object detection. What makes multi-view object instance detection difficult are strong changes in viewpoint, lighting conditions, high similarity of neighbouring objects, and strong variability in scale. By turning object detection and instance re-identification in different views into a joint learning task, we are able to incorporate both image appearance and geometric soft constraints into a single, multi-view detection process that is learnable end-to-end. We validate our method on a new, large data set of street-level panoramas of urban objects and show superior performance compared to various baselines. Our contribution is threefold: a large-scale, publicly available data set for multi-view instance detection and re-identification; an annotation tool custom-tailored for multi-view instance detection; and a novel, holistic multi-view instance detection and re-identification method that jointly models geometry and appearance across views.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Stacked Homography Transformations for Multi-View Pedestrian DetectionLiangchen Song, Jialian Wu, Ming Yang, Qian Zhang 等ICCV 2021 · 被引用 66 次
- The Auto Arborist Dataset: A Large-Scale Benchmark for Multiview Urban Forest Monitoring Under Domain ShiftSara Beery, Guanhang Wu, Trevor Edwards, Filip Pavetic 等CVPR 2022 · 被引用 53 次
- SAIL-VOS 3D: A Synthetic Dataset and Baselines for Object Detection and 3D Mesh Reconstruction From Video DataYuan-Ting Hu, Jiahong Wang, Raymond A. Yeh, Alexander G. SchwingCVPR 2021
相关 Paper
- Geometry-Aware Satellite-to-Ground Image Synthesis for Urban AreasXiaohu Lu, Zuoyue Li, Zhaopeng Cui, Martin R. Oswald 等CVPR 2020
- Viewpoint Equivariance for Multi-View 3D Object DetectionDian Chen, Jie Li, Vitor Guizilini, Rares Ambrus 等CVPR 2023
- Detection Based Part-level Articulated Object Reconstruction from Single RGBD ImageYuki Kawana, Tatsuya HaradaNeurIPS 2023 · 被引用 20 次
- Norm-Aware Embedding for Efficient Person SearchDi Chen, Shanshan Zhang, Jian Yang, Bernt SchieleCVPR 2020
- 3D-MOOD: Lifting 2D to 3D for Monocular Open-Set Object DetectionYung-Hsu Yang, Luigi Piccinelli, Mattia Segù, Siyuan Li 等ICCV 2025 · 被引用 2 次
