Equivariant Multi-View Networks
Carlos Esteves, Yinshuang Xu, Christine Allen-Blanchette, Kostas Daniilidis
Abstract
Several popular approaches to 3D vision tasks process multiple views of the input independently with deep neural networks pre-trained on natural images, achieving view permutation invariance through a single round of pooling over all views. We argue that this operation discards important information and leads to subpar global descriptors. In this paper, we propose a group convolutional approach to multiple view aggregation where convolutions are performed over a discrete subgroup of the rotation group, enabling, thus, joint reasoning over all views in an equivariant (instead of invariant) fashion, up to the very last layer. We further develop this idea to operate on smaller discrete homogeneous spaces of the rotation group, where a polar view representation is used to maintain equivariance with only a fraction of the number of input views. We set the new state of the art in several large scale 3D shape retrieval tasks, and show additional applications to panoramic scene classification. * Equal contribution.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5ee24486-709f-4c6c-ab06-9bcd9b287668Cited by top-tier papers38
- Vector Neurons: A General Framework for SO(3)-Equivariant NetworksCongyue Deng, Or Litany, Yueqi Duan, Adrien Poulenard et al.ICCV 2021 · 411 citations
- MVTN: Multi-View Transformation Network for 3D Shape RecognitionAbdullah Hamdi, Silvio Giancola, Bernard GhanemICCV 2021 · 280 citations
- A Group-Theoretic Framework for Data AugmentationShuxiao Chen, Edgar Dobriban, Jane H. LeeNeurIPS 2020 · 254 citations
- On Learning Sets of Symmetric ElementsHaggai Maron, Or Litany, Gal Chechik, Ethan FetayaICML 2020 · 148 citations
- You Only Hypothesize Once: Point Cloud Registration with Rotation-equivariant DescriptorsHaiping Wang, Yuan Liu, Zhen Dong, Wenping WangACM MM 2022 · 143 citations
Related papers
- E2PN: Efficient SE(3)-Equivariant Point NetworkMinghan Zhu, Maani Ghaffari, William A. Clark, Huei PengCVPR 2023
- Learning an Effective Equivariant 3D Descriptor Without SupervisionRiccardo Spezialetti, Samuele Salti, Luigi Di StefanoICCV 2019 · 41 citations
- View-GCN: View-Based Graph Convolutional Network for 3D Shape AnalysisXin Wei, Ruixuan Yu, Jian SunCVPR 2020
- Learning Rotation-Equivariant Features for Visual CorrespondenceJongmin Lee, Byungjin Kim, Seungwook Kim, Minsu ChoCVPR 2023
- SE(3) Equivariant Convolution and Transformer in Ray SpaceYinshuang Xu, Jiahui Lei, Kostas DaniilidisNeurIPS 2023 · 6 citations
