Towards Practical Control of Singular Values of Convolutional Layers
Alexandra Senderovich, Ekaterina Bulatova, Anton Obukhov, Maxim V. Rakhuba
Abstract
In general, convolutional neural networks (CNNs) are easy to train, but their essential properties, such as generalization error and adversarial robustness, are hard to control. Recent research demonstrated that singular values of convolutional layers significantly affect such elusive properties and offered several methods for controlling them. Nevertheless, these methods present an intractable computational challenge or resort to coarse approximations. In this paper, we offer a principled approach to alleviating constraints of the prior art at the expense of an insignificant reduction in layer expressivity. Our method is based on the tensor-train decomposition; it retains control over the actual singular values of convolutional mappings while providing structurally sparse and hardware-friendly representation. We demonstrate the improved properties of modern CNNs with our method and analyze its impact on the model performance, calibration, and adversarial robustness. The source code is available at: https://github.com/WhiteTeaDragon/practical_svd_conv
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 84e114ca-073f-459f-ad3d-a7b8ea0b0ae3Cited by top-tier papers4
- Training Scale-Invariant Neural Networks on the Sphere Can Happen in Three RegimesMaxim Kodryan, Ekaterina Lobacheva, Maksim Nakhodnov, Dmitry P. VetrovNeurIPS 2022 · 25 citations
- Training Robust Ensembles Requires Rethinking Lipschitz ContinuityAli Ebrahimpour Boroojeny, Hari Sundaram, Varun ChandrasekaranICLR 2025
- Can Spectral-Clipping Enable Better Learning While Forgetting Less for Low-Rank Adaptation?Hyowon Wi, Noseong ParkACL 2026
- An Adaptive Orthogonal Convolution Scheme for Efficient and Flexible CNN ArchitecturesThibaut Boissin, Franck Mamalet, Thomas Fel, Agustin Martin Picard et al.ICML 2025
Builds on9
- Reliable evaluation of adversarial robustness with an ensemble of diverse parameter-free attacksFrancesco Croce, Matthias HeinICML 2020 · 2,337 citations
- Spectral Regularization for Combating Mode Collapse in GANsKanglin Liu, Guoping Qiu, Wenming Tang, Fei ZhouICCV 2019 · 97 citations
- On the Practicality of Deterministic Epistemic UncertaintyJanis Postels, Mattia Segù, Tao Sun, Luca Daniel Sieber et al.ICML 2022 · 76 citations
- Skew Orthogonal ConvolutionsSahil Singla, Soheil FeiziICML 2021 · 76 citations
- TTOpt: A Maximum Volume Quantized Tensor Train-based Optimization and its Application to Reinforcement LearningKonstantin Sozykin, Andrei Chertkov, Roman Schutski, Anh-Huy Phan et al.NeurIPS 2022 · 62 citations
Related papers
- Transformed Low-Rank Parameterization Can Help Robust Generalization for Tensor Neural NetworksAndong Wang, Chao Li, Mingyuan Bai, Zhong Jin et al.NeurIPS 2023 · 12 citations
- Group-wise Inhibition based Feature Regularization for Robust ClassificationHaozhe Liu, Haoqian Wu, Weicheng Xie, Feng Liu et al.ICCV 2021 · 17 citations
- Generalized Depthwise-Separable Convolutions for Adversarially Robust and Efficient Neural NetworksHassan Dbouk, Naresh R. ShanbhagNeurIPS 2021 · 8 citations
- A Unified Weight Initialization Paradigm for Tensorial Convolutional Neural NetworksYu Pan, Zeyong Su, Ao Liu, Jingquan Wang et al.ICML 2022 · 15 citations
- Revisiting Sparse Convolutional Model for Visual RecognitionXili Dai, Mingyang Li, Pengyuan Zhai, Shengbang Tong et al.NeurIPS 2022 · 45 citations
