TESA: Tensor Element Self-Attention via Matricization
Francesca Babiloni, Ioannis Marras, Gregory G. Slabaugh, Stefanos Zafeiriou
摘要
Representation learning is a fundamental part of modern computer vision, where abstract representations of data are encoded as tensors optimized to solve problems like image segmentation and inpainting. Recently, self-attention in the form of a Non-Local Block has emerged as a powerful technique to enrich features, by capturing complex interdependencies in feature tensors. However, standard selfattention approaches leverage only spatial relationships, drawing similarities between vectors and overlooking correlations between channels. In this paper, we introduce a new method, called Tensor Element Self-Attention (TESA) that generalizes such work to capture interdependencies along all dimensions of the tensor using matricization. An order R tensor produces R results, one for each dimension. The results are then fused to produce an enriched output which encapsulates similarity among tensor elements. Additionally, we analyze self-attention mathematically, providing new perspectives on how it adjusts the singular values of the input feature tensor. With these new insights, we present experimental results demonstrating how TESA can benefit diverse problems including classification and instance segmentation. By simply adding a TESA module to existing networks, we substantially improve competitive baselines and set new state-of-the-art results for image inpainting on CelebA and low light raw-to-rgb image translation on SID.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Is Attention Better Than Matrix Decomposition?Zhengyang Geng, Meng-Hao Guo, Hongxu Chen, Xia Li 等ICLR 2021 · 被引用 171 次
- Fully Attentional Network for Semantic SegmentationQi Song, Jie Li, Chenghong Li, Hao Guo 等AAAI 2022 · 被引用 63 次
- Multilinear Mixture of Experts: Scalable Expert Specialization through FactorizationJames Oldfield, Markos Georgopoulos, Grigorios Chrysos, Christos Tzelepis 等NeurIPS 2024 · 被引用 41 次
- Poly-NL: Linear Complexity Non-local Layers With 3rd Order PolynomialsFrancesca Babiloni, Ioannis Marras, Filippos Kokkinos, Jiankang Deng 等ICCV 2021 · 被引用 14 次
它引用的顶会 Paper2
相关 Paper
- Steering Self-Supervised Feature Learning Beyond Local Pixel StatisticsSimon Jenni, Hailin Jin, Paolo FavaroCVPR 2020
- Attentive Normalization for Conditional Image GenerationYi Wang, Ying-Cong Chen, Xiangyu Zhang, Jian Sun 等CVPR 2020
- UCTGAN: Diverse Image Inpainting Based on Unsupervised Cross-Space TranslationLei Zhao, Qihang Mo, Sihuan Lin, Zhizhong Wang 等CVPR 2020
- Unifying Nonlocal Blocks for Neural NetworksLei Zhu, Qi She, Duo Li, Yanye Lu 等ICCV 2021 · 被引用 26 次
- TS-CAM: Token Semantic Coupled Attention Map for Weakly Supervised Object LocalizationWei Gao, Fang Wan, Xingjia Pan, Zhiliang Peng 等ICCV 2021 · 被引用 260 次
