Prototypical VoteNet for Few-Shot 3D Point Cloud Object Detection
Shizhen Zhao, Xiaojuan Qi
Abstract
Most existing 3D point cloud object detection approaches heavily rely on large amounts of labeled training data. However, the labeling process is costly and time-consuming. This paper considers few-shot 3D point cloud object detection, where only a few annotated samples of novel classes are needed with abundant samples of base classes. To this end, we propose Prototypical VoteNet to recognize and localize novel instances, which incorporates two new modules: Prototypical Vote Module (PVM) and Prototypical Head Module (PHM). Specifically, as the 3D basic geometric structures can be shared among categories, PVM is designed to leverage class-agnostic geometric prototypes, which are learned from base classes, to refine local features of novel categories. Then PHM is proposed to utilize class prototypes to enhance the global feature of each object, facilitating subsequent object localization and classification, which is trained by the episodic training strategy. To evaluate the model in this new setting, we contribute two new benchmark datasets, FS-ScanNet and FS-SUNRGBD. We conduct extensive experiments to demonstrate the effectiveness of Prototypical VoteNet, and our proposed method shows significant and consistent improvements compared to baselines on two benchmark datasets. This project will be available at https: //shizhen-zhao.github.io/FS3D_page/ .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers9
- Generalized Few-Shot Point Cloud Segmentation Via Geometric WordsYating Xu, Conghui Hu, Na Zhao, Gim Hee LeeICCV 2023 · 18 citations
- Commonsense Prototype for Outdoor Unsupervised 3D Object DetectionHai Wu, Shijia Zhao, Xun Huang, Chenglu Wen et al.CVPR 2024 · 15 citations
- Prototypical Variational Autoencoder for 3D Few-shot Object DetectionWeiliang Tang, Biqi Yang, Xianzhi Li, Yun-Hui Liu et al.NeurIPS 2023 · 8 citations
- Decompose Novel into Known: Part Concept Learning For 3D Novel Class DiscoveryTingyu Weng, Jun Xiao, Haiyong JiangNeurIPS 2023 · 6 citations
- Mixture-of-Scores: Robust Image-Text Data Valuation via Three Lines of CodeSitong Wu, Haoru Tan, Yukang Chen, Shaofeng Zhang et al.ICCV 2025 · 4 citations
Builds on29
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 1,467 citations
- Conditional Prompt Learning for Vision-Language ModelsKaiyang Zhou, Jingkang Yang, Chen Change Loy, Ziwei LiuCVPR 2022 · 1,438 citations
- Few-Shot Object Detection via Feature ReweightingBingyi Kang, Zhuang Liu, Xin Wang, Fisher Yu et al.ICCV 2019 · 835 citations
- Frustratingly Simple Few-Shot Object DetectionXin Wang, Thomas E. Huang, Joseph Gonzalez, Trevor Darrell et al.ICML 2020 · 723 citations
Related papers
- Boosting Few-shot 3D Point Cloud Segmentation via Query-Guided EnhancementZhenhua Ning, Zhuotao Tian, Guangming Lu, Wenjie PeiACM MM 2023 · 22 citations
- Few-Shot 3D Point Cloud Semantic SegmentationNa Zhao, Tat-Seng Chua, Gim Hee LeeCVPR 2021
- Few-Shot Incremental 3D Object Detection in Dynamic Indoor EnvironmentsYun Zhu, Jianjun Qian, Jian Yang, Jin Xie et al.CVPR 2026 · 2 citations
- ViewNet: A Novel Projection-Based Backbone with View Pooling for Few-shot Point Cloud ClassificationJiajing Chen, Minmin Yang, Senem VelipasalarCVPR 2023
- RBGNet: Ray-based Grouping for 3D Object DetectionHaiyang Wang, Shaoshuai Shi, Ze Yang, Rongyao Fang et al.CVPR 2022 · 63 citations
