Residual Relaxation for Multi-view Representation Learning
Yifei Wang, Zhengyang Geng, Feng Jiang, Chuming Li, Yisen Wang, Jiansheng Yang, Zhouchen Lin
Abstract
Multi-view methods learn representations by aligning multiple views of the same image and their performance largely depends on the choice of data augmentation. In this paper, we notice that some other useful augmentations, such as image rotation, are harmful for multi-view methods because they cause a semantic shift that is too large to be aligned well. This observation motivates us to relax the exact alignment objective to better cultivate stronger augmentations. Taking image rotation as a case study, we develop a generic approach, Pretext-aware Residual Relaxation (Prelax), that relaxes the exact alignment by allowing an adaptive residual vector between different views and encoding the semantic shift through pretext-aware learning. Extensive experiments on different backbones show that our method can not only improve multi-view methods with existing augmentations, but also benefit from stronger image augmentations like rotation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers23
- Chaos is a Ladder: A New Theoretical Understanding of Contrastive Learning via Augmentation OverlapYifei Wang, Qi Zhang, Yisen Wang, Jiansheng Yang et al.ICLR 2022 · 128 citations
- How Mask Matters: Towards Theoretical Understandings of Masked AutoencodersQi Zhang, Yifei Wang, Yisen WangNeurIPS 2022 · 119 citations
- Decoupled Self-supervised Learning for GraphsTeng Xiao, Zhengyu Chen, Zhimeng Guo, Zeyang Zhuang et al.NeurIPS 2022 · 75 citations
- A Novel Approach for Effective Multi-View Clustering with Information-Theoretic PerspectiveChenhang Cui, Yazhou Ren, Jingyu Pu, Jiawei Li et al.NeurIPS 2023 · 67 citations
- A Canonicalization Perspective on Invariant and Equivariant LearningGeorge Ma, Yifei Wang, Derek Lim, Stefanie Jegelka et al.NeurIPS 2024 · 38 citations
Builds on10
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- Big Self-Supervised Models are Strong Semi-Supervised LearnersTing Chen, Simon Kornblith, Kevin Swersky, Mohammad Norouzi et al.NeurIPS 2020 · 2,611 citations
- What Makes for Good Views for Contrastive Learning?Yonglong Tian, Chen Sun, Ben Poole, Dilip Krishnan et al.NeurIPS 2020 · 1,631 citations
- What Should Not Be Contrastive in Contrastive LearningTete Xiao, Xiaolong Wang, Alexei A. Efros, Trevor DarrellICLR 2021 · 338 citations
Related papers
- A Regularization-Guided Equivariant Approach for Image RestorationYulu Bai, Jiahong Fu, Qi Xie, Deyu MengCVPR 2025
- Equivariant Latent Alignment via Flow Matching under Group SymmetriesSunghyun Kim, Jaehoon Hahm, Jeongwoo Shin, Joonseok LeeICML 2026
- KeepAugment: A Simple Information-Preserving Data Augmentation ApproachChengyue Gong, Dilin Wang, Meng Li, Vikas Chandra et al.CVPR 2021
- Understanding the Role of Equivariance in Self-supervised LearningYifei Wang, Kaiwen Hu, Sharut Gupta, Ziyu Ye et al.NeurIPS 2024 · 10 citations
- A Flat Minima Perspective on Understanding Augmentations and Model RobustnessWeebum Yoo, Sung Whan YoonAAAI 2026 · 1 citation
