Prototypical Variational Autoencoder for 3D Few-shot Object Detection
Weiliang Tang, Biqi Yang, Xianzhi Li, Yun-Hui Liu, Pheng-Ann Heng, Chi-Wing Fu
摘要
Few-Shot 3D Point Cloud Object Detection (FS3D) is a challenging task, aiming to detect 3D objects of novel classes using only limited annotated samples for training. Considering that the detection performance highly relies on the quality of the latent features, we design a VAE-based prototype learning scheme, named prototypical VAE (P-VAE), to learn a probabilistic latent space for enhancing the diversity and distinctiveness of the sampled features. The network encodes a multi-center GMMlike posterior, in which each distribution centers at a prototype. For regularization, P-VAE incorporates a reconstruction task to preserve geometric information. To adopt P-VAE for the 3D object detection framework, we formulate Geometricinformative Prototypical VAE (GP-VAE) to handle varying geometric components and Class-specific Prototypical VAE (CP-VAE) to handle varying object categories. In the first stage, we harness GP-VAE to aid feature extraction from the input scene. In the second stage, we cluster the geometric-informative features into per-instance features and use CP-VAE to refine each instance feature with categorylevel guidance. Experimental results show the top performance of our approach over the state of the arts on two FS3D benchmarks. Quantitative ablations and qualitative prototype analysis further demonstrate that our probabilistic modeling can significantly boost prototype learning for FS3D.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Few-Shot Incremental 3D Object Detection in Dynamic Indoor EnvironmentsYun Zhu, Jianjun Qian, Jian Yang, Jin Xie 等CVPR 2026 · 被引用 2 次
- Promptable 3-D Object Localization with Latent Diffusion ModelsCheng-Yao Hong, Li-Heng Wang, Tyng-Luh LiuNeurIPS 2025
- From Dataset to Real-world: General 3D Object Detection via Generalized Cross-domain Few-shot LearningShuangzhi Li, Junlong Shen, Lei Ma, Xingyu LiAAAI 2026
它引用的顶会 Paper26
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 被引用 1,467 次
- PANet: Few-Shot Image Semantic Segmentation With Prototype AlignmentKaixin Wang, Jun Hao Liew, Yingtian Zou, Daquan Zhou 等ICCV 2019 · 被引用 1,404 次
- Few-Shot Object Detection via Feature ReweightingBingyi Kang, Zhuang Liu, Xin Wang, Fisher Yu 等ICCV 2019 · 被引用 835 次
- Frustratingly Simple Few-Shot Object DetectionXin Wang, Thomas E. Huang, Joseph Gonzalez, Trevor Darrell 等ICML 2020 · 被引用 723 次
- Prototypical Contrastive Learning of Unsupervised RepresentationsJunnan Li, Pan Zhou, Caiming Xiong, Steven C. H. HoiICLR 2021 · 被引用 484 次
相关 Paper
- Prototypical VoteNet for Few-Shot 3D Point Cloud Object DetectionShizhen Zhao, Xiaojuan QiNeurIPS 2022 · 被引用 33 次
- Few-Shot 3D Point Cloud Semantic SegmentationNa Zhao, Tat-Seng Chua, Gim Hee LeeCVPR 2021
- Boosting Few-shot 3D Point Cloud Segmentation via Query-Guided EnhancementZhenhua Ning, Zhuotao Tian, Guangming Lu, Wenjie PeiACM MM 2023 · 被引用 22 次
- Generalized Few-Shot Point Cloud Segmentation Via Geometric WordsYating Xu, Conghui Hu, Na Zhao, Gim Hee LeeICCV 2023 · 被引用 18 次
- Generating Point Cloud from Single Image in The Few Shot ScenarioYu Lin, Jinghui Guo, Yang Gao, Yi-Fan Li 等ACM MM 2021 · 被引用 6 次
