Spherization Layer: Representation Using Only Angles
Hoyong Kim, Kangil Kim
摘要
In neural network literature, angular similarity between feature vectors is frequently used for interpreting or re-using learned representations. However, the inner product in neural networks partially disperses information over the scales and angles of the involved input vectors and weight vectors. Therefore, when using only angular similarity on representations trained with the inner product, information loss occurs in downstream methods, which limits their performance. In this paper, we proposed the spherization layer to represent all information on angular similarity. The layer 1) maps the pre-activations of input vectors into the specific range of angles, 2) converts the angular coordinates of the vectors to Cartesian coordinates with an additional dimension, and 3) trains decision boundaries from hyperplanes, without bias parameters, passing through the origin. This approach guarantees that representation learning always occurs on the hyperspherical surface without the loss of any information unlike other projection-based methods. Furthermore, this method can be applied to any network by replacing an existing layer. We validate the functional correctness of the proposed method in a toy task, retention ability in well-known image classification tasks, and effectiveness in word analogy test and few-shot learning. Code is publicly available at https://github.com/GIST-IRR/spherization_layer * corresponding author 36th Conference on Neural Information Processing Systems (NeurIPS 2022).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper7
- Understanding Contrastive Representation Learning through Alignment and Uniformity on the HypersphereTongzhou Wang, Phillip IsolaICML 2020 · 被引用 2,360 次
- SphereFace2: Binary Classification is All You Need for Deep Face RecognitionYandong Wen, Weiyang Liu, Adrian Weller, Bhiksha Raj 等ICLR 2022 · 被引用 70 次
- Angular Visual HardnessBeidi Chen, Weiyang Liu, Zhiding Yu, Jan Kautz 等ICML 2020 · 被引用 57 次
- MMA Regularization: Decorrelating Weights of Neural Networks by Maximizing the Minimal AnglesZhennan Wang, Canqun Xiang, Wenbin Zou, Chen XuNeurIPS 2020 · 被引用 25 次
- CircleGAN: Generative Adversarial Learning across Spherical CirclesWoohyeon Shim, Minsu ChoNeurIPS 2020 · 被引用 12 次
相关 Paper
- PR Product: A Substitute for Inner Product in Neural NetworksZhennan Wang, Wenbin Zou, Chen XuICCV 2019 · 被引用 6 次
- Orthogonal Over-Parameterized TrainingWeiyang Liu, Rongmei Lin, Zhen Liu, James M. Rehg 等CVPR 2021
- An Analysis of SVD for Deep Rotation EstimationJake Levinson, Carlos Esteves, Kefan Chen, Noah Snavely 等NeurIPS 2020 · 被引用 131 次
- Rethinking Hebbian Principle: Low-Dimensional Structural Projection for Unsupervised LearningShikuang Deng, Jiayuan Zhang, Yuhang Wu, Ting Chen 等NeurIPS 2025 · 被引用 1 次
- OrthoRF: Exploring Orthogonality in Object-Centric RepresentationsDespoina Touska, Bastiaan Onne Fagginger Auer, Alexandru Onose, Tejaswi Kasarla 等ICLR 2026
