Efficient Equivariant Transfer Learning from Pretrained Models
Sourya Basu, Pulkit Katdare, Prasanna Sattigeri, Vijil Chenthamarakshan, Katherine Driggs-Campbell, Payel Das, Lav R. Varshney
摘要
Efficient transfer learning algorithms are key to the success of foundation models on diverse downstream tasks even with limited data. Recent works of Basu et al. (2023) and Kaba et al. (2022) propose group averaging (equitune) and optimization-based methods, respectively, over features from group-transformed inputs to obtain equivariant outputs from non-equivariant neural networks. While Kaba et al. (2022) are only concerned with training from scratch, we find that equitune performs poorly on equivariant zero-shot tasks despite good finetuning results. We hypothesize that this is because pretrained models provide better quality features for certain transformations than others and simply averaging them is deleterious. Hence, we propose -equitune that averages the features using importance weights, s. These weights are learned directly from the data using a small neural network, leading to excellent zero-shot and finetuned results that outperform equitune. Further, we prove that -equitune is equivariant and a universal approximator of equivariant functions. Additionally, we show that the method of Kaba et al. (2022) used with appropriate loss functions, which we call equizero, also gives excellent zero-shot and finetuned performance. Both equitune and equizero are special cases of -equitune. To show the simplicity and generality of our method, we validate on a wide range of diverse applications and models such as 1) image classification using CLIP, 2) deep Q-learning, 3) fairness in natural language generation (NLG), 4) compositional generalization in languages, and 5) image classification using pretrained CNNs such as Resnet and Alexnet.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Improving Equivariant Model Training via Constraint RelaxationStefanos Pertigkiozoglou, Evangelos Chatzipantazis, Shubhendu Trivedi, Kostas DaniilidisNeurIPS 2024 · 被引用 26 次
- Sample-specific Masks for Visual Reprogramming-based PromptingChengyi Cai, Zesheng Ye, Lei Feng, Jianzhong Qi 等ICML 2024 · 被引用 14 次
- Adaptive Canonicalization with Application to Invariant Anisotropic Geometric NetworksYa-Wei Eileen Lin, Ron LevieICLR 2026 · 被引用 4 次
- Geometric Embedding Alignment via Curvature Matching in Transfer LearningSung Moon Ko, Jaewan Lee, Sumin Lee, Soorin Yim 等ICML 2026
- Local Scale Equivariance with Latent Deep Equilibrium CanonicalizerMd Ashiqur Rahman, Chiao-An Yang, Michael N. Cheng, Lim Jun Hao 等ICCV 2025
它引用的顶会 Paper18
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- E(n) Equivariant Graph Neural NetworksVictor Garcia Satorras, Emiel Hoogeboom, Max WellingICML 2021 · 被引用 1,432 次
- A Practical Method for Constructing Equivariant Multilayer Perceptrons for Arbitrary Matrix GroupsMarc Finzi, Max Welling, Andrew Gordon WilsonICML 2021 · 被引用 226 次
- MDP Homomorphic Networks: Group Symmetries in Reinforcement LearningElise van der Pol, Daniel E. Worrall, Herke van Hoof, Frans A. Oliehoek 等NeurIPS 2020 · 被引用 203 次
相关 Paper
- Equi-Tuning: Group Equivariant Fine-Tuning of Pretrained ModelsSourya Basu, Prasanna Sattigeri, Karthikeyan Natesan Ramamurthy, Vijil Chenthamarakshan 等AAAI 2023 · 被引用 25 次
- Robust fine-tuning of zero-shot modelsMitchell Wortsman, Gabriel Ilharco, Jong Wook Kim, Mike Li 等CVPR 2022 · 被引用 364 次
- Group Equivariant Conditional Neural ProcessesMakoto Kawano, Wataru Kumagai, Akiyoshi Sannai, Yusuke Iwasawa 等ICLR 2021 · 被引用 22 次
- A Closer Look at the Few-Shot Adaptation of Large Vision-Language ModelsJulio Silva-Rodríguez, Sina Hajimiri, Ismail Ben Ayed, Jose DolzCVPR 2024 · 被引用 32 次
- Generalized Logit Adjustment: Calibrating Fine-tuned Models by Removing Label Bias in Foundation ModelsBeier Zhu, Kaihua Tang, Qianru Sun, Hanwang ZhangNeurIPS 2023 · 被引用 50 次
