SVHAN: Sequential View Based Hierarchical Attention Network for 3D Shape Recognition
Yue Zhao, Weizhi Nie, An-An Liu, Zan Gao, Yuting Su
Abstract
As an important field of multimedia, 3D shape recognition has attracted much research attention in recent years. A lot of deep learning models have been proposed for effective 3D shape representation. The view-based methods show the superiority due to the comprehensive exploration of the visual characteristics with the help of established 2D CNN architectures. Generally, the current approaches contain the following disadvantages: First, the most majority of methods lack the consideration for sequential information among the multiple views, which can provide descriptive characteristics for shape representation. Second, the incomprehensive exploration for the multi-view correlations directly affects the discrimination of shape descriptors. Finally, roughly aggregating multi-view features leads to the loss of descriptive information, which limits the shape representation effectiveness. To handle these issues, we propose a novel sequential view based hierarchical attention network (SVHAN) for 3D shape recognition. Specifically, we first divide the view sequence into several view blocks. Then, we introduce a novel hierarchical feature aggregation module (HFAM), which hierarchically exploits the view-level, block-level, and shape-level features, the intra- and inter- view-block correlations are also captured to improve the discrimination of learned features. Subsequently, a novel selective fusion module (SFM) is designed for feature aggregation, considering the correlations between different levels and preserving effective information. Finally, discriminative and informative shape descriptors are generated for the recognition task. We validate the effectiveness of our proposed method on two public databases. The experimental results show the superiority of SVHAN against the current state-of-the-art approaches.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 1c52dbe2-0b80-4ebf-a3dc-6e262a04d218Related papers
- View-GCN: View-Based Graph Convolutional Network for 3D Shape AnalysisXin Wei, Ruixuan Yu, Jian SunCVPR 2020
- HMTN: Hierarchical Multi-scale Transformer Network for 3D Shape RecognitionYue Zhao, Weizhi Nie, Zan Gao, Anan LiuACM MM 2022 · 3 citations
- Hierarchical View Predictor: Unsupervised 3D Global Feature Learning through Hierarchical Prediction among Unordered ViewsZhizhong Han, Xiyang Wang, Yu-Shen Liu, Matthias ZwickerACM MM 2021 · 9 citations
- Pairwise View Weighted Graph Network for View-based 3D Model RetrievalZan Gao, Yin-Ming Li, Weili Guan, Weizhi Nie et al.SIGIR 2020 · 9 citations
- Learning Feature Aggregation for Deep 3D Morphable ModelsZhixiang Chen, Tae-Kyun KimCVPR 2021
