Equivariant Adaptation of Large Pretrained Models
Arnab Kumar Mondal, Siba Smarak Panigrahi, Oumar Kaba, Sai Mudumba, Siamak Ravanbakhsh
摘要
Equivariant networks are specifically designed to ensure consistent behavior with respect to a set of input transformations, leading to higher sample efficiency and more accurate and robust predictions. However, redesigning each component of prevalent deep neural network architectures to achieve chosen equivariance is a difficult problem and can result in a computationally expensive network during both training and inference. A recently proposed alternative towards equivariance that removes the architectural constraints is to use a simple canonicalization network that transforms the input to a canonical form before feeding it to an unconstrained prediction network. We show here that this approach can effectively be used to make a large pretrained network equivariant. However, we observe that the produced canonical orientations can be misaligned with those of the training distribution, hindering performance. Using dataset-dependent priors to inform the canonicalization function, we are able to make large pretrained models equivariant while maintaining their performance. This significantly improves the robustness of these models to deterministic transformations of the data, such as rotations. We believe this equivariant adaptation of large pretrained models can help their domain-specific applications with known symmetry priors.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper27
- Equivariant Frames and the Impossibility of Continuous CanonicalizationNadav Dym, Hannah Lawrence, Jonathan W. SiegelICML 2024 · 被引用 38 次
- Improving Equivariant Model Training via Constraint RelaxationStefanos Pertigkiozoglou, Evangelos Chatzipantazis, Shubhendu Trivedi, Kostas DaniilidisNeurIPS 2024 · 被引用 26 次
- Equivariance via Minimal Frame Averaging for More Symmetries and EfficiencyYuchao Lin, Jacob Helwig, Shurui Gui, Shuiwang JiICML 2024 · 被引用 20 次
- A Generative Model of Symmetry TransformationsJames Urquhart Allingham, Bruno Mlodozeniec, Shreyas Padhy, Javier Antorán 等NeurIPS 2024 · 被引用 16 次
- Sample-specific Masks for Visual Reprogramming-based PromptingChengyi Cai, Zesheng Ye, Lei Feng, Jianzhong Qi 等ICML 2024 · 被引用 14 次
它引用的顶会 Paper16
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Big Self-Supervised Models are Strong Semi-Supervised LearnersTing Chen, Simon Kornblith, Kevin Swersky, Mohammad Norouzi 等NeurIPS 2020 · 被引用 2,611 次
- Scaling Vision TransformersXiaohua Zhai, Alexander Kolesnikov, Neil Houlsby, Lucas BeyerCVPR 2022 · 被引用 767 次
相关 Paper
- Equivariance with Learned Canonicalization FunctionsSékou-Oumar Kaba, Arnab Kumar Mondal, Yan Zhang, Yoshua Bengio 等ICML 2023 · 被引用 109 次
- Adaptive Canonicalization with Application to Invariant Anisotropic Geometric NetworksYa-Wei Eileen Lin, Ron LevieICLR 2026 · 被引用 4 次
- Towards Diffeomorphism-Equivariant Neural Networks via CanonicalizationJosephine Elisabeth Oettinger, Zakhar Shumaylov, Johannes Bostelmann, Jan Lellmann 等ICML 2026
- Learning (Approximately) Equivariant Networks via Constrained OptimizationAndrei Manolache, Luiz F. O. Chamon, Mathias NiepertNeurIPS 2025 · 被引用 12 次
- Normalization Equivariance for Arbitrary Backbones, with Application to Image DenoisingYoussef Saied, François FleuretICML 2026
