REViT: Roto-reflection Equivariant Convolutional Vision Transformer
Sheir A. Zaheer, Alexander Holston, Chan Youn Park
摘要
In this paper, we propose a discrete roto-reflection group equivariant vision transformer with convolutional attention. Roto-reflection equivariant networks preserve the rotational, flip and positional symmetry in feature maps, making them useful for tasks where orientation of the inputs is relevant to the model outputs. In image classification and object detection, most of the studies on roto-reflection equivariant models have focused on using convolutional neural networks rather than vision transformers. In this paper, we examine the challenges involved in achieving equivariance in vision transformers, and we propose a simpler way to implement a discretized roto-reflection group equivariant vision transformer. The experimental results demonstrate that our approach outperforms the existing approaches for developing discrete roto-reflection group equivariant neural networks for image classification.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper19
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- CvT: Introducing Convolutions to Vision TransformersHaiping Wu, Bin Xiao, Noel Codella, Mengchen Liu 等ICCV 2021 · 被引用 2,397 次
- E(n) Equivariant Graph Neural NetworksVictor Garcia Satorras, Emiel Hoogeboom, Max WellingICML 2021 · 被引用 1,432 次
- Equivariant message passing for the prediction of tensorial properties and molecular spectraKristof Schütt, Oliver T. Unke, Michael GasteggerICML 2021 · 被引用 736 次
- Rethinking and Improving Relative Position Encoding for Vision TransformerKan Wu, Houwen Peng, Minghao Chen, Jianlong Fu 等ICCV 2021 · 被引用 427 次
相关 Paper
- Co-Attentive Equivariant Neural Networks: Focusing Equivariance On Transformations Co-Occurring in DataDavid W. Romero, Mark HoogendoornICLR 2020 · 被引用 24 次
- Group Equivariant Stand-Alone Self-Attention For VisionDavid W. Romero, Jean-Baptiste CordonnierICLR 2021 · 被引用 72 次
- Reflection and Rotation Symmetry Detection via Equivariant LearningAhyun Seo, Byungjin Kim, Suha Kwak, Minsu ChoCVPR 2022 · 被引用 12 次
- Learning Partial Equivariances From DataDavid W. Romero, Suhas LohitNeurIPS 2022 · 被引用 54 次
- Attentive Group Equivariant Convolutional NetworksDavid W. Romero, Erik J. Bekkers, Jakub M. Tomczak, Mark HoogendoornICML 2020 · 被引用 99 次
