GO-N3RDet: Geometry Optimized NeRF-enhanced 3D Object Detector
Zechuan Li, Hongshan Yu, Yihao Ding, Jinhao Qiao, Basim Azam, Naveed Akhtar
摘要
We propose GO-N3RDet, a scene-geometry optimized multi-view 3D object detector enhanced by neural radiance fields. The key to accurate 3D object detection is in effective voxel representation. However, due to occlusion and lack of 3D information, constructing 3D features from multi-view 2D images is challenging. Addressing that, we introduce a unique 3D positional information embedded voxel optimization mechanism to fuse multi-view features. To prioritize neural field reconstruction in object regions, we also devise a double importance sampling scheme for the NeRF branch of our detector. We additionally propose an opacity optimization module for precise voxel opacity prediction by enforcing multi-view consistency constraints. Moreover, to further improve voxel density consistency across multiple perspectives, we incorporate ray distance as a weighting factor to minimize cumulative ray errors. Our unique modules synergetically form an end-to-end neural model that establishes new state-of-the-art in NeRF-based multi-view 3D detection, verified with extensive experiments on ScanNet and ARKITScenes. Code will be available at https://github.com/ZechuanLi/GO-N3RDet.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper18
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 被引用 1,467 次
- SASA: Semantics-Augmented Set Abstraction for Point-Based 3D Object DetectionChen Chen, Zhe Chen, Jing Zhang, Dacheng TaoAAAI 2022 · 被引用 166 次
- NeRF-Det: Learning Geometry-Aware Volumetric Representation for Multi-View 3D Object DetectionChenfeng Xu, Bichen Wu, Ji Hou, Sam S. Tsai 等ICCV 2023 · 被引用 71 次
- VENet: Voting Enhancement Network for 3D Object DetectionQian Xie, Yu-Kun Lai, Jing Wu, Zhoutao Wang 等ICCV 2021 · 被引用 60 次
- HUGS: Holistic Urban 3D Scene Understanding via Gaussian SplattingHongyu Zhou, Jiahao Shao, Lu Xu, Dongfeng Bai 等CVPR 2024 · 被引用 50 次
相关 Paper
- To View Transform or Not to View Transform: NeRF-based Pre-training PerspectiveHyeonjun Jeong, Juyeb Shin, Dongsuk KumICLR 2026 · 被引用 2 次
- MVSDet: Multi-View Indoor 3D Object Detection via Efficient Plane SweepsYating Xu, Chen Li, Gim Hee LeeNeurIPS 2024 · 被引用 10 次
- NeRF-RPN: A general framework for object detection in NeRFsBenran Hu, Junkai Huang, Yichen Liu, Yu-Wing Tai 等CVPR 2023
- CVT-xRF: Contrastive In-Voxel Transformer for 3D Consistent Radiance Fields from Sparse InputsYingji Zhong, Lanqing Hong, Zhenguo Li, Dan XuCVPR 2024 · 被引用 4 次
- OcRFDet: Object-Centric Radiance Fields for Multi-View 3D Object Detection in Autonomous DrivingMingqian Ji, Shanshan Zhang, Jian YangICCV 2025 · 被引用 2 次
