Pseudo-Stereo for Monocular 3D Object Detection in Autonomous Driving
Yi-Nan Chen, Hang Dai, Yong Ding
摘要
Pseudo-LiDAR 3D detectors have made remarkable progress in monocular 3D detection by enhancing the capability of perceiving depth with depth estimation networks, and using LiDAR-based 3D detection architectures. The Advanced stereo 3D detectors can also accurately localize 3D objects. The gap in image-to-image generation for stereo views is much smaller than that in image-to-LiDAR generation. Motivated by this, we propose a Pseudo-Stereo 3D detection framework with three novel virtual view generation methods, including image-level generation, feature-level generation, and feature-clone, for detecting 3D objects from a single image. Our analysis of depth-aware learning shows that the depth loss is effective in only feature-level virtual view generation and the estimated depth map is effective in both image-level and feature-level in our framework. We propose a disparity-wise dynamic convolution with dynamic kernels sampled from the disparity feature map to filter the features adaptively from a single image for generating virtual image features, which eases the feature degradation caused by the depth estimation errors. Till submission (November 18, 2021), our Pseudo-Stereo 3D detection framework ranks 1 st on car, pedestrian, and cyclist among the monocular 3D detectors with publications on the KITTI-3D benchmark. The code is released at https://github.com/revisitq/Pseudo-Stereo-3D.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- MonoNeRD: NeRF-like Representations for Monocular 3D Object DetectionJunkai Xu, Liang Peng, Haoran Chen, Hao Li 等ICCV 2023 · 被引用 54 次
- Mono3DVG: 3D Visual Grounding in Monocular ImagesYang Zhan, Yuan Yuan, Zhitong XiongAAAI 2024 · 被引用 38 次
- Attention-Based Depth Distillation with 3D-Aware Positional Encoding for Monocular 3D Object DetectionZizhang Wu, Yunzhe Wu, Jian Pu, Xianzhi Li 等AAAI 2023 · 被引用 29 次
- Learning Occupancy for Monocular 3D Object DetectionLiang Peng, Junkai Xu, Haoran Cheng, Zheng Yang 等CVPR 2024 · 被引用 21 次
- MonoDiff: Monocular 3D Object Detection and Pose Estimation with Diffusion ModelsYasiru Ranasinghe, Deepti Hegde, Vishal M. PatelCVPR 2024 · 被引用 21 次
它引用的顶会 Paper25
- M3D-RPN: Monocular 3D Region Proposal Network for Object DetectionGarrick Brazil, Xiaoming LiuICCV 2019 · 被引用 542 次
- Disentangling Monocular 3D Object DetectionAndrea Simonelli, Samuel Rota Bulò, Lorenzo Porzi, Manuel Lopez-Antequera 等ICCV 2019 · 被引用 504 次
- Pseudo-LiDAR++: Accurate Depth for 3D Object Detection in Autonomous DrivingYurong You, Yan Wang, Wei-Lun Chao, Divyansh Garg 等ICLR 2020 · 被引用 439 次
- Is Pseudo-Lidar needed for Monocular 3D Object detection?Dennis Park, Rares Ambrus, Vitor Guizilini, Jie Li 等ICCV 2021 · 被引用 404 次
- Accurate Monocular 3D Object Detection via Color-Embedded 3D Reconstruction for Autonomous DrivingXinzhu Ma, Zhihui Wang, Haojie Li, Pengbo Zhang 等ICCV 2019 · 被引用 339 次
相关 Paper
- Learning Depth-Guided Convolutions for Monocular 3D Object DetectionMingyu Ding, Yuqi Huo, Hongwei Yi, Zhe Wang 等CVPR 2020
- Neighbor-Vote: Improving Monocular 3D Object Detection through Neighbor Distance VotingXiaomeng Chu, Jiajun Deng, Yao Li, Zhenxun Yuan 等ACM MM 2021 · 被引用 24 次
- End-to-End Pseudo-LiDAR for Image-Based 3D Object DetectionRui Qian, Divyansh Garg, Yan Wang, Yurong You 等CVPR 2020
- IDA-3D: Instance-Depth-Aware 3D Object Detection From Stereo Vision for Autonomous DrivingWanli Peng, Hao Pan, He Liu, Yi SunCVPR 2020
- Disp R-CNN: Stereo 3D Object Detection via Shape Prior Guided Instance Disparity EstimationJiaming Sun, Linghao Chen, Yiming Xie, Siyu Zhang 等CVPR 2020
