DisARM: Displacement Aware Relation Module for 3D Detection
Yao Duan, Chenyang Zhu, Yuqing Lan, Renjiao Yi, Xinwang Liu, Kai Xu
Abstract
We introduce Displacement Aware Relation Module (DisARM), a novel neural network module for enhancing the performance of 3D object detection in point cloud scenes. The core idea is extracting the most principal contextual information is critical for detection while the target is incomplete or featureless. We find that relations between proposals provide a good representation to describe the context. However, adopting relations between all the object or patch proposals for detection is inefficient, and an imbalanced combination of local and global relations brings extra noise that could mislead the training. Rather than working with all relations, we find that training with relations only between the most representative ones, or an-chors, can significantly boost the detection performance. Good anchors should be semantic-aware with no ambiguity and able to describe the whole layout of a scene with no redundancy. To find the anchors, we first perform a preliminary relation anchor module with an objectness-aware sampling approach and then devise a displacement based module for weighing the relation importance for better utilization of contextual information. This lightweight relation module leads to significantly higher accuracy of object instance detection when being plugged into the state-of-the-art detectors. Evaluations on the public benchmarks of real-world scenes show that our method achieves the state-of-the-art performance on both SUN RGB-D and Scan-Net V2. The code and models are publicly available at https://github.com/YaraDuan/DisARM.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e4049602-a5e1-488e-ab03-088965672f91Cited by top-tier papers1
Ask how each one uses itBuilds on11
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 1,467 citations
- Group-Free 3D Object Detection via TransformersZe Liu, Zheng Zhang, Yue Cao, Han Hu et al.ICCV 2021 · 368 citations
- A Hierarchical Graph Network for 3D Object Detection on Point CloudsJintai Chen, Biwen Lei, Qingyu Song, Haochao Ying et al.CVPR 2020
- Learning 3D Semantic Scene Graphs From 3D Indoor ReconstructionsJohanna Wald, Helisa Dhamo, Nassir Navab, Federico TombariCVPR 2020
- ImVoteNet: Boosting 3D Object Detection in Point Clouds With Image VotesCharles R. Qi, Xinlei Chen, Or Litany, Leonidas J. GuibasCVPR 2020
Related papers
- SPGroup3D: Superpoint Grouping Network for Indoor 3D Object DetectionYun Zhu, Le Hui, Yaqi Shen, Jin XieAAAI 2024 · 24 citations
- Correlation Field for Boosting 3D Object Detection in Structured ScenesJianhua Sun, Haoshu Fang, Xianghui Zhu, Jiefeng Li et al.AAAI 2022 · 8 citations
- Density-Based Clustering for 3D Object Detection in Point CloudsSyeda Mariam Ahmed, Chee-Meng ChewCVPR 2020
- CAGroup3D: Class-Aware Grouping for 3D Object Detection on Point CloudsHaiyang Wang, Lihe Ding, Shaocong Dong, Shaoshuai Shi et al.NeurIPS 2022 · 110 citations
- Real-time 3D Object Detection with Inference-Aligned LearningChenyu Zhao, Xianwei Zheng, Zimin Xia, Linwei Yue et al.AAAI 2026 · 1 citation
