Self-Supervised Learning with Kernel Dependence Maximization
Yazhe Li, Roman Pogodin, Danica J. Sutherland, Arthur Gretton
摘要
We approach self-supervised learning of image representations from a statistical dependence perspective, proposing Self-Supervised Learning with the Hilbert-Schmidt Independence Criterion (SSL-HSIC). SSL-HSIC maximizes dependence between representations of transformations of an image and the image identity, while minimizing the kernelized variance of those representations. This framework yields a new understanding of InfoNCE, a variational lower bound on the mutual information (MI) between different transformations. While the MI itself is known to have pathologies which can result in learning meaningless representations, its bound is much better behaved: we show that it implicitly approximates SSL-HSIC (with a slightly different regularizer). Our approach also gives us insight into BYOL, a negative-free SSL method, since SSL-HSIC similarly learns local neighborhoods of samples. SSL-HSIC allows us to directly optimize statistical dependence in time linear in the batch size, without restrictive data assumptions or indirect mutual information estimators. Trained with or without a target network, SSL-HSIC matches the current state-of-the-art for standard linear evaluation on ImageNet [1], semi-supervised learning and transfer to other classification and vision tasks such as semantic segmentation, depth estimation and object recognition. Code is available at https://github.com/deepmind/ssl_hsic .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper34
- Understanding Negative Samples in Instance Discriminative Self-supervised Representation LearningKento Nozawa, Issei SatoNeurIPS 2021 · 被引用 56 次
- Self-Supervised Learning with an Information Maximization CriterionSerdar Ozsoy, Shadi Hamdan, Sercan Ö. Arik, Deniz Yuret 等NeurIPS 2022 · 被引用 54 次
- Quality-Agnostic Deepfake Detection with Intra-model Collaborative LearningBinh Minh Le, Simon S. WooICCV 2023 · 被引用 50 次
- Exploring the Gap between Collapsed & Whitened Features in Self-Supervised LearningBobby He, Mete OzayICML 2022 · 被引用 31 次
- On the Surrogate Gap between Contrastive and Supervised LossesHan Bao, Yoshihiro Nagano, Kento NozawaICML 2022 · 被引用 27 次
它引用的顶会 Paper18
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun 等ICML 2021 · 被引用 2,942 次
相关 Paper
- Mean Shift for Self-Supervised LearningSoroush Abbasi Koohpayegani, Ajinkya Tejankar, Hamed PirsiavashICCV 2021 · 被引用 104 次
- Making Sense of Dependence: Efficient Black-box Explanations Using Dependence MeasurePaul Novello, Thomas Fel, David VigourouxNeurIPS 2022 · 被引用 48 次
- S4L: Self-Supervised Semi-Supervised LearningLucas Beyer, Xiaohua Zhai, Avital Oliver, Alexander KolesnikovICCV 2019 · 被引用 854 次
- Self-supervised learning with rotation-invariant kernelsLéon Zheng, Gilles Puy, Elisa Riccietti, Patrick Pérez 等ICLR 2023 · 被引用 3 次
- Contrastive-Equivariant Self-Supervised Learning Improves Alignment with Primate Visual Area ITThomas E. Yerxa, Jenelle Feather, Eero P. Simoncelli, SueYeon ChungNeurIPS 2024 · 被引用 12 次
