Controllable Orthogonalization in Training DNNs
Lei Huang, Li Liu, Fan Zhu, Diwen Wan, Zehuan Yuan, Bo Li, Ling Shao
摘要
Orthogonality is widely used for training deep neural networks (DNNs) due to its ability to maintain all singular values of the Jacobian close to 1 and reduce redundancy in representation. This paper proposes a computationally efficient and numerically stable orthogonalization method using Newton's iteration (ONI), to learn a layer-wise orthogonal weight matrix in DNNs. ONI works by iteratively stretching the singular values of a weight matrix towards 1. This property enables it to control the orthogonality of a weight matrix by its number of iterations. We show that our method improves the performance of image classification networks by effectively controlling the orthogonality to provide an optimal tradeoff between optimization benefits and representational capacity reduction. We also show that ONI stabilizes the training of generative adversarial networks (GANs) by maintaining the Lipschitz continuity of a network, similar to spectral normalization (SN), and further outperforms SN by providing controllable orthogonality.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper26
- A Dynamical System Perspective for Lipschitz Neural NetworksLaurent Meunier, Blaise Delattre, Alexandre Araujo, Alexandre AllauzenICML 2022 · 被引用 69 次
- Deep Isometric Learning for Visual RecognitionHaozhi Qi, Chong You, Xiaolong Wang, Yi Ma 等ICML 2020 · 被引用 57 次
- projUNN: efficient method for training deep networks with unitary matricesBobak Toussi Kiani, Randall Balestriero, Yann LeCun, Seth LloydNeurIPS 2022 · 被引用 42 次
- LOT: Layer-wise Orthogonal Training on Improving l2 Certified RobustnessXiaojun Xu, Linyi Li, Bo LiNeurIPS 2022 · 被引用 42 次
- Orthogonal Graph Neural NetworksKai Guo, Kaixiong Zhou, Xia Hu, Yu Li 等AAAI 2022 · 被引用 41 次
相关 Paper
- Convolutional Normalization: Improving Deep Convolutional Network Robustness and TrainingSheng Liu, Xiao Li, Yuexiang Zhai, Chong You 等NeurIPS 2021 · 被引用 30 次
- Why Spectral Normalization Stabilizes GANs: Analysis and ImprovementsZinan Lin, Vyas Sekar, Giulia FantiNeurIPS 2021 · 被引用 67 次
- Stochastic Whitening Batch NormalizationShengdong Zhang, Ehsan Nezhadarya, Homa Fashandi, Jiayi Liu 等CVPR 2021
- Stable Rank Normalization for Improved Generalization in Neural Networks and GANsAmartya Sanyal, Philip H. S. Torr, Puneet K. DokaniaICLR 2020 · 被引用 58 次
- Efficient Bound of Lipschitz Constant for Convolutional Layers by Gram IterationBlaise Delattre, Quentin Barthélemy, Alexandre Araujo, Alexandre AllauzenICML 2023 · 被引用 20 次
