Bridging Mini-Batch and Asymptotic Analysis in Contrastive Learning: From InfoNCE to Kernel-Based Losses
Panagiotis Koromilas, Giorgos Bouritsas, Theodoros Giannakopoulos, Mihalis Nicolaou, Yannis Panagakis
Abstract
What do different contrastive learning (CL) losses actually optimize for? Although multiple CL methods have demonstrated remarkable representation learning capabilities, the differences in their inner workings remain largely opaque. In this work, we analyse several CL families and prove that, under certain conditions, they admit the same minimisers when optimizing either their batch-level objectives or their expectations asymptotically. In both cases, an intimate connection with the hyperspherical energy minimisation (HEM) problem resurfaces. Drawing inspiration from this, we introduce a novel CL objective, coined Decoupled Hyperspherical Energy Loss (DHEL). DHEL simplifies the problem by decoupling the target hyperspherical energy from the alignment of positive examples while preserving the same theoretical guarantees. Going one step further, we show the same results hold for another relevant CL family, namely kernel contrastive learning (KCL), with the additional advantage of the expected loss being independent of batch size, thus identifying the minimisers in the non-asymptotic regime. Empirical results demonstrate improved downstream performance and robustness across combinations of different batch sizes and hyperparameters and reduced dimensionality collapse, on several computer vision datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c8f287cd-457d-4dee-89e7-745bb5a1d433Cited by top-tier papers3
- SNAPHARD CONTRAST LEARNINGChangpu Meng, Jie Yang, Wanqing Li, Yi GuoICLR 2026
- Supervised Contrastive Learning from Weakly-Labeled Audio Segments for Musical Version MatchingJoan Serrà, Recep Oguz Araz, Dmitry Bogdanov, Yuki MitsufujiICML 2025
- Neural Collapse by Design: Learning Class Prototypes on the HyperspherePanagiotis Koromilas, Theodoros Giannakopoulos, Mihalis Nicolaou, Yannis PanagakisICML 2026
Builds on24
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Understanding Contrastive Representation Learning through Alignment and Uniformity on the HypersphereTongzhou Wang, Phillip IsolaICML 2020 · 2,360 citations
- Contrastive Learning with Hard Negative SamplesJoshua David Robinson, Ching-Yao Chuang, Suvrit Sra, Stefanie JegelkaICLR 2021 · 999 citations
- With a Little Help from My Friends: Nearest-Neighbor Contrastive Learning of Visual RepresentationsDebidatta Dwibedi, Yusuf Aytar, Jonathan Tompson, Pierre Sermanet et al.ICCV 2021 · 542 citations
- Understanding Dimensional Collapse in Contrastive Self-supervised LearningLi Jing, Pascal Vincent, Yann LeCun, Yuandong TianICLR 2022 · 467 citations
Related papers
- Provable Discriminative Hyperspherical Embedding for Out-of-Distribution DetectionZhipeng Zou, Sheng Wan, Guangyu Li, Bo Han et al.AAAI 2025 · 1 citation
- How to Exploit Hyperspherical Embeddings for Out-of-Distribution Detection?Yifei Ming, Yiyou Sun, Ousmane Dia, Yixuan LiICLR 2023 · 24 citations
- Dissecting Supervised Constrastive LearningFlorian Graf, Christoph D. Hofer, Marc Niethammer, Roland KwittICML 2021 · 73 citations
- Hyperspherical Consistency RegularizationCheng Tan, Zhangyang Gao, Lirong Wu, Siyuan Li et al.CVPR 2022 · 16 citations
- I-Con: A Unifying Framework for Representation LearningShaden Naif Alshammari, John R. Hershey, Axel Feldmann, William T. Freeman et al.ICLR 2025
