A Novel Patch Convolutional Neural Network for View-based 3D Model Retrieval
Zan Gao, Yuxiang Shao, Weili Guan, Meng Liu, Zhiyong Cheng, Shengyong Chen
Abstract
In industrial enterprises, effective retrieval of three-dimensional (3-D) computer-aided design (CAD) models can greatly save time and cost in new product development and manufacturing, thus, many researchers have focused on it. Recently, many view-based 3D model retrieval methods have been proposed and have achieved state-of-the-art performance. However, most of these methods focus on extracting more discriminative view-level features and effectively aggregating the multi-view images of a 3D model, and the latent relationship among these multi-view images is not fully explored. Thus, we tackle this problem from the perspective of exploiting the relationships between patch features to capture long-range associations among multi-view images. To capture associations among views, in this work, we propose a novel patch convolutional neural network (PCNN ) for view-based 3D model retrieval. Specifically, we first employ a CNN to extract patch features of each view image separately. Second, a novel neural network module named PatchConv is designed to exploit intrinsic relationships between neighboring patches in the feature space to capture long-range associations among multi-view images. Then, an adaptive weighted view layer is further embedded into PCNN to automatically assign a weight to each view according to the similarity between each view feature and the view-pooling feature. Finally, a discrimination loss function is employed to extract the discriminative 3D model feature, which consists of softmax loss values generated by the fusion classifier and the specific classifier. Extensive experimental results on two public 3D model retrieval benchmarks, namely, the ModelNet40, and ModelNet10, demonstrate that our proposed PCNN can outperform state-of-the-art approaches, with mAP values of 93.67%, and 96.23%, respectively.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e7472c27-d6ca-41eb-9a07-ec869a9fd9c8Builds on5
- Deconfounded Video Moment Retrieval with Causal InterventionXun Yang, Fuli Feng, Wei Ji, Meng Wang et al.SIGIR 2021 · 198 citations
- Tree-Augmented Cross-Modal Encoding for Complex-Query Video RetrievalXun Yang, Jianfeng Dong, Yixin Cao, Xun Wang et al.SIGIR 2020 · 131 citations
- View N-Gram Network for 3D Object RetrievalXinwei He, Tengteng Huang, Song Bai, Xiang BaiICCV 2019 · 65 citations
- Semantic Consistency Guided Instance Feature Alignment for 2D Image-Based 3D Shape RetrievalHeyu Zhou, Weizhi Nie, Dan Song, Nian Hu et al.ACM MM 2020 · 27 citations
- Multi-graph Convolutional Network for Unsupervised 3D Shape RetrievalWeizhi Nie, Yue Zhao, An-An Liu, Zan Gao et al.ACM MM 2020 · 10 citations
Related papers
- Pairwise View Weighted Graph Network for View-based 3D Model RetrievalZan Gao, Yin-Ming Li, Weili Guan, Weizhi Nie et al.SIGIR 2020 · 9 citations
- Learning Relationships for Multi-View 3D Object RecognitionZe Yang, Liwei WangICCV 2019 · 166 citations
- View-GCN: View-Based Graph Convolutional Network for 3D Shape AnalysisXin Wei, Ruixuan Yu, Jian SunCVPR 2020
- Enhancing 2D Representation via Adjacent Views for 3D Shape RetrievalCheng Xu, Zhaoqun Li, Qiang Qiu, Biao Leng et al.ICCV 2019 · 19 citations
- MVTN: Multi-View Transformation Network for 3D Shape RecognitionAbdullah Hamdi, Silvio Giancola, Bernard GhanemICCV 2021 · 280 citations
