ViewNet: A Novel Projection-Based Backbone with View Pooling for Few-shot Point Cloud Classification
Jiajing Chen, Minmin Yang, Senem Velipasalar
Abstract
Although different approaches have been proposed for 3D point cloud-related tasks, few-shot learning (FSL) of 3D point clouds still remains under-explored. In FSL, unlike traditional supervised learning, the classes of training and test data do not overlap, and a model needs to recognize unseen classes from only a few samples. Existing FSL methods for 3D point clouds employ point-based models as their backbone. Yet, based on our extensive experiments and analysis, we first show that using a point-based backbone is not the most suitable FSL approach, since (i) a large number of points' features are discarded by the max pooling operation used in 3D point-based backbones, decreasing the ability of representing shape information;
(ii) point-based backbones are sensitive to occlusion. To address these issues, we propose employing a projectionand 2D Convolutional Neural Network-based backbone, referred to as the ViewNet, for FSL from 3D point clouds.
Our approach first projects a 3D point cloud onto six different views to alleviate the issue of missing points. Also, to generate more descriptive and distinguishing features, we propose View Pooling, which combines different projected plane combinations into five groups and performs maxpooling on each of them. The experiments performed on the ModelNet40, ScanObjectNN and ModelNet40-C datasets, with cross validation, show that our method consistently outperforms the state-of-the-art baselines. Moreover, compared to traditional image classification backbones, such as ResNet, the proposed ViewNet can extract more distinguishing features from multiple views of a point cloud. We also show that ViewNet can be used as a backbone with different FSL heads and provides improved performance compared to traditionally used backbones.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6c1bf760-eb83-47d1-b3bb-8038d864ba5eBuilds on6
- Revisiting Point Cloud Classification: A New Benchmark Dataset and Classification Model on Real-World DataMikaela Angelina Uy, Quang-Hieu Pham, Binh-Son Hua, Duc Thanh Nguyen et al.ICCV 2019 · 1,003 citations
- Walk in the Cloud: Learning Curves for Point Clouds Shape AnalysisTiange Xiang, Chaoyi Zhang, Yang Song, Jianhui Yu et al.ICCV 2021 · 369 citations
- Revisiting Point Cloud Shape Classification with a Simple and Effective BaselineAnkit Goyal, Hei Law, Bowei Liu, Alejandro Newell et al.ICML 2021 · 297 citations
- Learning Geometry-Disentangled Representation for Complementary Understanding of 3D Object Point CloudMutian Xu, Junhao Zhang, Zhipeng Zhou, Mingye Xu et al.AAAI 2021 · 175 citations
- Why Discard if You can Recycle?: A Recycling Max Pooling Module for 3D Point Cloud AnalysisJiajing Chen, Burak Kakillioglu, Huantao Ren, Senem VelipasalarCVPR 2022 · 20 citations
Related papers
- Few-Shot 3D Point Cloud Semantic Segmentation via Stratified Class-Specific Attention Based Transformer NetworkCanyu Zhang, Zhenyao Wu, Xinyi Wu, Ziyu Zhao et al.AAAI 2023 · 31 citations
- Prototypical VoteNet for Few-Shot 3D Point Cloud Object DetectionShizhen Zhao, Xiaojuan QiNeurIPS 2022 · 33 citations
- PointCLIP: Point Cloud Understanding by CLIPRenrui Zhang, Ziyu Guo, Wei Zhang, Kunchang Li et al.CVPR 2022
- Generating Point Cloud from Single Image in The Few Shot ScenarioYu Lin, Jinghui Guo, Yang Gao, Yi-Fan Li et al.ACM MM 2021 · 6 citations
- Voint Cloud: Multi-View Point Cloud Representation for 3D UnderstandingAbdullah Hamdi, Silvio Giancola, Bernard GhanemICLR 2023 · 4 citations
