Poly-View Contrastive Learning
Amitis Shidani, R. Devon Hjelm, Jason Ramapuram, Russell Webb, Eeshan Gunesh Dhekane, Dan Busbridge
Abstract
Contrastive learning typically matches pairs of related views among a number of unrelated negative views. Views can be generated (e.g. by augmentations) or be observed. We investigate matching when there are more than two related views which we call poly-view tasks, and derive new representation learning objectives using information maximization and sufficient statistics. We show that with unlimited computation, one should maximize the number of related views, and with a fixed compute budget, it is beneficial to decrease the number of unique samples whilst increasing the number of views of those samples. In particular, poly-view contrastive models trained for 128 epochs with batch size 256 outperform Sim-CLR trained for 1024 epochs at batch size 4096 on ImageNet1k, challenging the belief that contrastive models require large batch sizes and many training epochs.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- What Variables Affect Out-of-Distribution Generalization in Pretrained Models?Md Yousuf Harun, Kyungbok Lee, Gianmarco J. Gallardo, Giri Krishnan et al.NeurIPS 2024 · 16 citations
- Contrasting Multiple Representations with the Multi-Marginal Matching GapZoe Piran, Michal Klein, James Thornton, Marco CuturiICML 2024 · 9 citations
- FedAugment: Table Augmentation Search over Decentralized Data RepositoriesLennart Behme, Emil Badura, Leonard Geißler, Matthias Böhm et al.VLDB 2026
Builds on23
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
Related papers
- X-Sample Contrastive Loss: Improving Contrastive Learning with Sample Similarity GraphsVlad Sobal, Mark Ibrahim, Randall Balestriero, Vivien Cabannes et al.ICLR 2025
- AUC-CL: A Batchsize-Robust Framework for Self-Supervised Contrastive Representation LearningRohan Sharma, Kaiyi Ji, Zhiqiang Xu, Changyou ChenICLR 2024 · 5 citations
- EASEMVC: Efficient Dual Selection Mechanism for Deep Multi-View ClusteringBaili Xiao, Zhibin Dong, Ke Liang, Suyuan Liu et al.CVPR 2025
- A Mutual Information Perspective on Federated Contrastive LearningChristos Louizos, Matthias Reisser, Denis KorzhenkovICLR 2024 · 7 citations
- Decomposed Mutual Information Estimation for Contrastive Representation LearningAlessandro Sordoni, Nouha Dziri, Hannes Schulz, Geoffrey J. Gordon et al.ICML 2021 · 37 citations
