Analyzing Data-Centric Properties for Graph Contrastive Learning
Puja Trivedi, Ekdeep Singh Lubana, Mark Heimann, Danai Koutra, Jayaraman J. Thiagarajan
摘要
Recent analyses of self-supervised learning (SSL) find the following data-centric properties to be critical for learning good representations: invariance to task-irrelevant semantics, separability of classes in some latent space, and recoverability of labels from augmented samples. However, given their discrete, non-Euclidean nature, graph datasets and graph SSL methods are unlikely to satisfy these properties. This raises the question: how do graph SSL methods, such as contrastive learning (CL), work well? To systematically probe this question, we perform a generalization analysis for CL when using generic graph augmentations (GGAs), with a focus on data-centric properties. Our analysis yields formal insights into the limitations of GGAs and the necessity of task-relevant augmentations. As we empirically show, GGAs do not induce task-relevant invariances on common benchmark datasets, leading to only marginal gains over naive, untrained baselines. Our theory motivates a synthetic data generation process that enables control over task-relevant information and boasts pre-defined optimal augmentations. This flexible benchmark helps us identify yet unrecognized limitations in advanced augmentation techniques (e.g., automated methods). Overall, our work rigorously contextualizes, both empirically and theoretically, the effects of data-centric properties on augmentation strategies and learning paradigms for graph SSL.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Simple and Asymmetric Graph Contrastive Learning without AugmentationsTeng Xiao, Huaisheng Zhu, Zhengyu Chen, Suhang WangNeurIPS 2023 · 被引用 86 次
- Data-Centric Learning from Unlabeled Graphs with Diffusion ModelGang Liu, Eric Inae, Tong Zhao, Jiaxin Xu 等NeurIPS 2023 · 被引用 32 次
- Better with Less: A Data-Active Perspective on Pre-Training Graph Neural NetworksJiarong Xu, Renhong Huang, Xin Jiang, Yuxuan Cao 等NeurIPS 2023 · 被引用 26 次
- Architecture Matters: Uncovering Implicit Mechanisms in Graph Contrastive LearningXiaojun Guo, Yifei Wang, Zeming Wei, Yisen WangNeurIPS 2023 · 被引用 20 次
- Adversarial Training for Graph Neural Networks: Pitfalls, Solutions, and New DirectionsLukas Gosch, Simon Geisler, Daniel Sturm, Bertrand Charpentier 等NeurIPS 2023 · 被引用 19 次
它引用的顶会 Paper39
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
- Graph Contrastive Learning with AugmentationsYuning You, Tianlong Chen, Yongduo Sui, Ting Chen 等NeurIPS 2020 · 被引用 3,042 次
相关 Paper
- Boosting Graph Contrastive Learning via Graph Contrastive SaliencyChunyu Wei, Yu Wang, Bing Bai, Kai Ni 等ICML 2023 · 被引用 31 次
- SGCL: Semantic-aware Graph Contrastive Learning with Lipschitz Graph AugmentationJinhao Cui, Heyan Chai, Xu Yang, Ye Ding 等ICDE 2024 · 被引用 1 次
- Label-invariant Augmentation for Semi-Supervised Graph ClassificationHan Yue, Chunhui Zhang, Chuxu Zhang, Hongfu LiuNeurIPS 2022 · 被引用 37 次
- Augmentations in Graph Contrastive Learning: Current Methodological Flaws & Towards Better PracticesPuja Trivedi, Ekdeep Singh Lubana, Yujun Yan, Yaoqing Yang 等WWW 2022 · 被引用 59 次
- Spectral Augmentation for Self-Supervised Learning on GraphsLu Lin, Jinghui Chen, Hongning WangICLR 2023 · 被引用 16 次
