OrthCaps: An Orthogonal CapsNet with Sparse Attention Routing and Pruning
Xinyu Geng, Jiaming Wang, Jiawei Gong, Yuerong Xue, Jun Xu, Fanglin Chen, Xiaolin Huang
Abstract
Redundancy is a persistent challenge in Capsule Networks (CapsNet), leading to high computational costs and parameter counts. Although previous studies have introduced pruning after the initial capsule layer, dynamic routing's fully connected nature and non-orthogonal weight matrices reintroduce redundancy in deeper layers. Besides, dynamic routing requires iterating to converge, further increasing computational demands. In this paper, we propose an Orthogonal Capsule Network (OrthCaps) to reduce redundancy, improve routing performance and decrease parameter counts. Firstly, an efficient pruned capsule layer is introduced to discard redundant capsules. Secondly, dynamic routing is replaced with orthogonal sparse attention routing, eliminating the need for iterations and fully connected structures. Lastly, weight matrices during routing are orthogonalized to sustain low capsule similarity, which is the first approach to use Householder orthogonal decomposition to enforce orthogonality in Cap-sNet. Our experiments on baseline datasets affirm the efficiency and robustness of OrthCaps in classification tasks, in which ablation studies validate the criticality of each component. OrthCaps-Shallow outperforms other Capsule Network benchmarks on four datasets, utilizing only 110k parameters -a mere 1.25% of a standard Capsule Network's total. To the best of our knowledge, it achieves the smallest parameter count among existing Capsule Networks. Similarly, OrthCaps-Deep demonstrates competitive performance across four datasets, utilizing only 1.2% of the parameters required by its counterparts.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 31b98240-77b0-4d32-93c0-58cfc6cda2e8Cited by top-tier papers2
- MSPCaps: A Multi-Scale Patchify Capsule Network with Cross-Agreement Routing for Visual RecognitionYudong Hu, Yueju Han, Rui Sun, Jinke RenAAAI 2026
- ParseCaps: An Interpretable Parsing Capsule Network for Medical Image DiagnosisXinyu Geng, Jiaming Wang, Xiaolin Huang, Fanglin Chen et al.AAAI 2025
Builds on10
- Efficient Riemannian Optimization on the Stiefel Manifold via the Cayley TransformJun Li, Fuxin Li, Sinisa TodorovicICLR 2020 · 139 citations
- Orthogonalizing Convolutional Layers with the Cayley TransformAsher Trockman, J. Zico KolterICLR 2021 · 137 citations
- Capsules with Inverted Dot-Product Attention RoutingYao-Hung Hubert Tsai, Nitish Srivastava, Hanlin Goh, Ruslan SalakhutdinovICLR 2020 · 91 citations
- Skew Orthogonal ConvolutionsSahil Singla, Soheil FeiziICML 2021 · 76 citations
- Deep Isometric Learning for Visual RecognitionHaozhi Qi, Chong You, Xiaolong Wang, Yi Ma et al.ICML 2020 · 57 citations
Related papers
- Capsule Network Is Not More Robust Than Convolutional NetworkJindong Gu, Volker Tresp, Han HuCVPR 2021
- Hyperbolic Capsule Networks for Multi-Label ClassificationBoli Chen, Xin Huang, Lin Xiao, Liping JingACL 2020 · 20 citations
- A Receptor Skeleton for Capsule Neural NetworksJintai Chen, Hongyun Yu, Chengde Qian, Danny Z. Chen et al.ICML 2021 · 5 citations
- Improving the Robustness of Capsule Networks to Image Affine TransformationsJindong Gu, Volker TrespCVPR 2020
- Towards an Effective Orthogonal Dictionary Convolution StrategyYishi Li, Kunran Xu, Rui Lai, Lin GuAAAI 2022 · 4 citations
