Hierarchical Direction Perception via Atomic Dot-Product Operators for Rotation-Invariant Point Clouds Learning
Chenyu Hu, Xiaotong Li, Hao Zhu, Biao Hou
Abstract
Point cloud processing has become a cornerstone technology in many 3D vision tasks. However, arbitrary rotations introduce variations in point cloud orientations, posing a longstanding challenge for effective representation learning. The core of this issue is the disruption of the point cloud's intrinsic directional characteristics caused by rotational perturbations. Recent methods attempt to implicitly model rotational equivariance and invariance, preserving directional information and propagating it into deep semantic spaces. Yet, they often fall short of fully exploiting the multiscale directional nature of point clouds to enhance feature representations. To address this, we propose the Direction-Perceptive Vector Network (DiPVNet). At its core is an atomic dotproduct operator that simultaneously encodes directional selectivity and rotation invariance-endowing the network with both rotational symmetry modeling and adaptive directional perception. At the local level, we introduce a Learnable Local Dot-Product (L2DP) Operator, which enables interactions between a center point and its neighbors to adaptively capture the non-uniform local structures of point clouds. At the global level, we leverage generalized harmonic analysis to prove that the dot-product between point clouds and spherical sampling vectors is equivalent to a direction-aware spherical Fourier transform (DASFT). This leads to the construction of a global directional response spectrum for modeling holistic directional structures. We rigorously prove the rotation invariance of both operators. Extensive experiments on challenging scenarios involving noise and large-angle rotations demonstrate that DiPVNet achieves state-of-the-art performance on point cloud classification and segmentation tasks. Our code is available at https://github.com/wxszreal0/DiPVNet .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on22
- E(n) Equivariant Graph Neural NetworksVictor Garcia Satorras, Emiel Hoogeboom, Max WellingICML 2021 · 1,432 citations
- Revisiting Point Cloud Classification: A New Benchmark Dataset and Classification Model on Real-World DataMikaela Angelina Uy, Quang-Hieu Pham, Binh-Son Hua, Duc Thanh Nguyen et al.ICCV 2019 · 1,003 citations
- Equivariant message passing for the prediction of tensorial properties and molecular spectraKristof Schütt, Oliver T. Unke, Michael GasteggerICML 2021 · 736 citations
- Learning from Protein Structure with Geometric Vector PerceptronsBowen Jing, Stephan Eismann, Patricia Suriana, Raphael John Lamarre Townshend et al.ICLR 2021 · 627 citations
- Vector Neurons: A General Framework for SO(3)-Equivariant NetworksCongyue Deng, Or Litany, Yueqi Duan, Adrien Poulenard et al.ICCV 2021 · 411 citations
Related papers
- Pointwise Rotation-Invariant Network with Adaptive Sampling and 3D Spherical Voxel ConvolutionYang You, Yujing Lou, Qi Liu, Yu-Wing Tai et al.AAAI 2020 · 73 citations
- PaRot: Patch-Wise Rotation-Invariant Network via Feature Disentanglement and Pose RestorationDingxin Zhang, Jianhui Yu, Chaoyi Zhang, Weidong CaiAAAI 2023 · 17 citations
- SCF-Net: Learning Spatial Contextual Features for Large-Scale Point Cloud SegmentationSiqi Fan, Qiulei Dong, Fenghua Zhu, Yisheng Lv et al.CVPR 2021
- QPoint: End-to-End Lightweight Point Cloud Processing via Robust Quaternion Feature LearningZhouzhiming Zhou, Yong He, Chaoxu Mu, Qiaoyun Wu et al.ICML 2026
- Rotationally Equivariant 3D Object DetectionHong-Xing Yu, Jiajun Wu, Li YiCVPR 2022 · 31 citations
