3D Video Object Detection with Learnable Object-Centric Global Optimization
Jiawei He, Yuntao Chen, Naiyan Wang, Zhaoxiang Zhang
摘要
We explore long-term temporal visual correspondencebased optimization for 3D video object detection in this work. Visual correspondence refers to one-to-one mappings for pixels across multiple images. Correspondencebased optimization is the cornerstone for 3D scene reconstruction but is less studied in 3D video object detection, because moving objects violate multi-view geometry constraints and are treated as outliers during scene reconstruction. We address this issue by treating objects as first-class citizens during correspondence-based optimization. In this work, we propose BA-Det, an end-to-end optimizable object detector with object-centric temporal correspondence learning and featuremetric object bundle adjustment. Empirically, we verify the effectiveness and efficiency of BA-Det for multiple baseline 3D detectors under various setups. Our BA-Det achieves SOTA performance on the large-scale Waymo Open Dataset (WOD) with only marginal computation cost. Our code is available at https://github.com/jiaweihe1996/BA-Det .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper18
- M3D-RPN: Monocular 3D Region Proposal Network for Object DetectionGarrick Brazil, Xiaoming LiuICCV 2019 · 被引用 542 次
- PETRv2: A Unified Framework for 3D Perception from Multi-Camera ImagesYingfei Liu, Junjie Yan, Fan Jia, Shuailin Li 等ICCV 2023 · 被引用 513 次
- Geometry Uncertainty Projection Network for Monocular 3D Object DetectionYan Lu, Xinzhu Ma, Lei Yang, Tianzhu Zhang 等ICCV 2021 · 被引用 294 次
- Pixel-Perfect Structure-from-Motion with Featuremetric RefinementPhilipp Lindenberger, Paul-Edouard Sarlin, Viktor Larsson, Marc PollefeysICCV 2021 · 被引用 266 次
- LIGA-Stereo: Learning LiDAR Geometry Aware Representations for Stereo-based 3D DetectorXiaoyang Guo, Shaoshuai Shi, Xiaogang Wang, Hongsheng LiICCV 2021 · 被引用 132 次
相关 Paper
- DetZero: Rethinking Offboard 3D Object Detection with Long-term Sequential Point CloudsTao Ma, Xuemeng Yang, Hongbin Zhou, Xin Li 等ICCV 2023 · 被引用 46 次
- 3D-MAN: 3D Multi-Frame Attention Network for Object DetectionZetong Yang, Yin Zhou, Zhifeng Chen, Jiquan NgiamCVPR 2021
- R3D3: Dense 3D Reconstruction of Dynamic Scenes from Multiple CamerasAron Schmied, Tobias Fischer, Martin Danelljan, Marc Pollefeys 等ICCV 2023 · 被引用 52 次
- Point2Seq: Detecting 3D Objects as SequencesYujing Xue, Jiageng Mao, Minzhe Niu, Hang Xu 等CVPR 2022 · 被引用 24 次
- Joint Spatial-Temporal Optimization for Stereo 3D Object TrackingPeiliang Li, Jieqi Shi, Shaojie ShenCVPR 2020
