Augmentation Component Analysis: Modeling Similarity via the Augmentation Overlaps
Lu Han, Han-Jia Ye, De-Chuan Zhan
摘要
Self-supervised learning aims to learn a embedding space where semantically similar samples are close. Contrastive learning methods pull views of samples together and push different samples away, which utilizes semantic invariance of augmentation but ignores the relationship between samples. To better exploit the power of augmentation, we observe that semantically similar samples are more likely to have similar augmented views. Therefore, we can take the augmented views as a special description of a sample. In this paper, we model such a description as the augmentation distribution and we call it augmentation feature. The similarity in augmentation feature reflects how much the views of two samples overlap and is related to their semantical similarity. Without computational burdens to explicitly estimate values of the augmentation feature, we propose Augmentation Component Analysis (ACA) with a contrastive-like loss to learn principal components and an on-the-fly projection loss to embed data. ACA equals an efficient dimension reduction by PCA and extracts low-dimensional embeddings, theoretically preserving the similarity of augmentation distribution between samples. Empirical results show our method can achieve competitive results against various traditional contrastive learning methods on different benchmarks. Code available at https://github.com/hanlu-nju/AugCA .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- When Can We Approximate Wide Contrastive Models with Neural Tangent Kernels and Principal Component Analysis?Gautham Govind Anil, Pascal Mattia Esser, Debarghya GhoshdastidarAAAI 2025 · 被引用 1 次
- Learning Without Augmenting: Unsupervised Time Series Representation Learning via Frame ProjectionsBerken Utku Demirel, Christian HolzNeurIPS 2025 · 被引用 1 次
它引用的顶会 Paper22
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun 等ICML 2021 · 被引用 2,942 次
- Big Self-Supervised Models are Strong Semi-Supervised LearnersTing Chen, Simon Kornblith, Kevin Swersky, Mohammad Norouzi 等NeurIPS 2020 · 被引用 2,611 次
相关 Paper
- Contrastive Learning Can Find An Optimal Basis For Approximately View-Invariant FunctionsDaniel D. Johnson, Ayoub El Hanchi, Chris J. MaddisonICLR 2023 · 被引用 1 次
- Revisiting Contrastive Learning through the Lens of Neighborhood Component Analysis: an Integrated FrameworkChing-Yun Ko, Jeet Mohapatra, Sijia Liu, Pin-Yu Chen 等ICML 2022 · 被引用 15 次
- An Augmentation-Aware Theory for Self-Supervised Contrastive LearningJingyi Cui, Hongwei Wen, Yisen WangICML 2025
- Intriguing Properties of Contrastive LossesTing Chen, Calvin Luo, Lala LiNeurIPS 2021 · 被引用 206 次
- Identity-Disentangled Adversarial Augmentation for Self-supervised LearningKaiwen Yang, Tianyi Zhou, Xinmei Tian, Dacheng TaoICML 2022 · 被引用 13 次
