PR Product: A Substitute for Inner Product in Neural Networks
Zhennan Wang, Wenbin Zou, Chen Xu
Abstract
In this paper, we analyze the inner product of weight vector w and data vector x in neural networks from the perspective of vector orthogonal decomposition and prove that the direction gradient of w decreases with the angle between them close to 0 or π. We propose the Projection and Rejection Product (PR Product) to make the direction gradient of w independent of the angle and consistently larger than the one in standard inner product while keeping the forward propagation identical. As a reliable substitute for standard inner product, the PR Product can be applied into many existing deep learning modules, so we develop the PR Product version of fully connected layer, convolutional layer and LSTM layer. In static image classification, the experiments on CIFAR10 and CIFAR100 datasets demonstrate that the PR Product can robustly enhance the ability of various state-of-the-art classification networks. On the task of image captioning, even without any bells and whistles, our PR Product version of captioning model can compete or outperform the state-of-the-art models on MS COCO dataset. Code has been made available at: https: //github.com/wzn0828/PR_Product .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cd03fba2-ece9-45a3-a69a-713afe0614f4Cited by top-tier papers1
- Spherization Layer: Representation Using Only AnglesHoyong Kim, Kangil KimNeurIPS 2022 · 3 citations
Related papers
- Feature Projection for Improved Text ClassificationQi Qin, Wenpeng Hu, Bing LiuACL 2020 · 66 citations
- Quaternion Product Units for Deep Learning on 3D Rotation GroupsXuan Zhang, Shaofei Qin, Yi Xu, Hongteng XuCVPR 2020
- Reversible Column NetworksYuxuan Cai, Yizhuang Zhou, Qi Han, Jianjian Sun et al.ICLR 2023 · 21 citations
- Normalized and Geometry-Aware Self-Attention Network for Image CaptioningLongteng Guo, Jing Liu, Xinxin Zhu, Peng Yao et al.CVPR 2020
- Is normalization indispensable for training deep neural network?Jie Shao, Kai Hu, Changhu Wang, Xiangyang Xue et al.NeurIPS 2020 · 70 citations
