Diversity Matters: Fully Exploiting Depth Clues for Reliable Monocular 3D Object Detection
Zhuoling Li, Zhan Qu, Yang Zhou, Jianzhuang Liu, Haoqian Wang, Lihui Jiang
摘要
As an inherently ill-posed problem, depth estimation from single images is the most challenging part of monocular 3D object detection (M3OD). Many existing methods rely on preconceived assumptions to bridge the missing spatial information in monocular images, and predict a sole depth value for every object of interest. However, these assumptions do not always hold in practical applications. To tackle this problem, we propose a depth solving system that fully explores the visual clues from the subtasks in M3OD and generates multiple estimations for the depth of each target. Since the depth estimations rely on different assumptions in essence, they present diverse distributions. Even if some assumptions collapse, the estimations established on the remaining assumptions are still reliable. In addition, we develop a depth selection and combination strategy. This strategy is able to remove abnormal estimations caused by collapsed assumptions, and adaptively combine the remaining estimations into a single one. In this way, our depth solving system becomes more precise and robust. Exploiting the clues from multiple subtasks of M3OD and without introducing any extra information, our method surpasses the current best method by more than 20% relatively on the Moderate level of test split in the KITTI 3D object detection benchmark, while still maintaining real-time efficiency. * Zhuoling Li and Zhan Qu contributed equally. This work was done when Zhuoling Li was an intern at Huawei Noah's Ark Lab. † Corresponding author. (a) Direct estimation (1 depth): 0.6220 (b) Depth from height (3 depths): 0.5831 (c) Depth from keypoint (16 depths): 0.3210 (d) Our system (20 depths): 0.2819
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper22
- Time Will Tell: New Outlooks and A Baseline for Temporal Multi-View 3D Object DetectionJinhyung Park, Chenfeng Xu, Shijia Yang, Kurt Keutzer 等ICLR 2023 · 被引用 71 次
- MonoUNI: A Unified Vehicle and Infrastructure-side Monocular 3D Object Detection Network with Sufficient Depth CluesJinrang Jia, Zhenjia Li, Yifeng ShiNeurIPS 2023 · 被引用 69 次
- MonoCD: Monocular 3D Object Detection with Complementary DepthsLongfei Yan, Pei Yan, Shengzhou Xiong, Xuanyu Xiang 等CVPR 2024 · 被引用 52 次
- Learning Occupancy for Monocular 3D Object DetectionLiang Peng, Junkai Xu, Haoran Cheng, Zheng Yang 等CVPR 2024 · 被引用 21 次
- Decoupled Pseudo-Labeling for Semi-Supervised Monocular 3D Object DetectionJiacheng Zhang, Jiaming Li, Xiangru Lin, Wei Zhang 等CVPR 2024 · 被引用 17 次
它引用的顶会 Paper21
- M3D-RPN: Monocular 3D Region Proposal Network for Object DetectionGarrick Brazil, Xiaoming LiuICCV 2019 · 被引用 542 次
- Disentangling Monocular 3D Object DetectionAndrea Simonelli, Samuel Rota Bulò, Lorenzo Porzi, Manuel Lopez-Antequera 等ICCV 2019 · 被引用 504 次
- Gaussian YOLOv3: An Accurate and Fast Object Detector Using Localization Uncertainty for Autonomous DrivingJiwoong Choi, Dayoung Chun, Hyun Kim, Hyuk-Jae LeeICCV 2019 · 被引用 445 次
- Is Pseudo-Lidar needed for Monocular 3D Object detection?Dennis Park, Rares Ambrus, Vitor Guizilini, Jie Li 等ICCV 2021 · 被引用 404 次
- Accurate Monocular 3D Object Detection via Color-Embedded 3D Reconstruction for Autonomous DrivingXinzhu Ma, Zhihui Wang, Haojie Li, Pengbo Zhang 等ICCV 2019 · 被引用 339 次
相关 Paper
- Objects Are Different: Flexible Monocular 3D Object DetectionYunpeng Zhang, Jiwen Lu, Jie ZhouCVPR 2021
- Monocular 3D Object Detection with Decoupled Structured Polygon Estimation and Height-Guided Depth EstimationYingjie Cai, Buyu Li, Zeyu Jiao, Hongsheng Li 等AAAI 2020 · 被引用 100 次
- MoGDE: Boosting Mobile Monocular 3D Object Detection with Ground Depth EstimationYunsong Zhou, Quan Liu, Hongzi Zhu, Yunzhe Li 等NeurIPS 2022 · 被引用 23 次
- MonoGround: Detecting Monocular 3D Objects from the GroundZequn Qin, Xi LiCVPR 2022 · 被引用 68 次
- Depth-Conditioned Dynamic Message Propagation for Monocular 3D Object DetectionLi Wang, Liang Du, Xiaoqing Ye, Yanwei Fu 等CVPR 2021
