Learning Canonical View Representation for 3D Shape Recognition with Arbitrary Views
Xin Wei, Yifei Gong, Fudong Wang, Xing Sun, Jian Sun
摘要
In this paper, we focus on recognizing 3D shapes from arbitrary views, i.e., arbitrary numbers and positions of viewpoints. It is a challenging and realistic setting for view-based 3D shape recognition. We propose a canonical view representation to tackle this challenge. We first transform the original features of arbitrary views to a fixed number of view features, dubbed canonical view representation, by aligning the arbitrary view features to a set of learnable reference view features using optimal transport. In this way, each 3D shape with arbitrary views is represented by a fixed number of canonical view features, which are further aggregated to generate a rich and robust 3D shape representation for shape recognition. We also propose a canonical view feature separation constraint to enforce that the view features in canonical view representation can be embedded into scattered points in a Euclidean space. Experiments on the ModelNet40, ScanObjectNN, and RGBD datasets show that our method achieves competitive results under the fixed viewpoint settings, and significantly outperforms the applicable methods under the arbitrary view setting.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Learning Dense Object Descriptors from Multiple Views for Low-shot Category GeneralizationStefan Stojanov, Anh Thai, Zixuan Huang, James M. RehgNeurIPS 2022 · 被引用 6 次
- Dual Pose-invariant Embeddings: Learning Category and Object-specific Discriminative Representations for Recognition and RetrievalRohan Sarkar, Avinash C. KakCVPR 2024 · 被引用 3 次
- Incremental 3D Semantic Scene Graph Prediction from RGB SequencesShun-Cheng Wu, Keisuke Tateno, Nassir Navab, Federico TombariCVPR 2023
它引用的顶会 Paper11
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without ConvolutionsWenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan 等ICCV 2021 · 被引用 4,909 次
- Revisiting Point Cloud Classification: A New Benchmark Dataset and Classification Model on Real-World DataMikaela Angelina Uy, Quang-Hieu Pham, Binh-Son Hua, Duc Thanh Nguyen 等ICCV 2019 · 被引用 1,003 次
相关 Paper
- MVTN: Multi-View Transformation Network for 3D Shape RecognitionAbdullah Hamdi, Silvio Giancola, Bernard GhanemICCV 2021 · 被引用 280 次
- View-GCN: View-Based Graph Convolutional Network for 3D Shape AnalysisXin Wei, Ruixuan Yu, Jian SunCVPR 2020
- Novel Object Viewpoint Estimation Through Reconstruction AlignmentMohamed El Banani, Jason J. Corso, David F. FouheyCVPR 2020
- Voint Cloud: Multi-View Point Cloud Representation for 3D UnderstandingAbdullah Hamdi, Silvio Giancola, Bernard GhanemICLR 2023 · 被引用 4 次
- Recognizing Objects From Any View With Object and Viewer-Centered RepresentationsSainan Liu, Vincent Nguyen, Isaac Rehg, Zhuowen TuCVPR 2020
