Augmentations in Graph Contrastive Learning: Current Methodological Flaws & Towards Better Practices
Puja Trivedi, Ekdeep Singh Lubana, Yujun Yan, Yaoqing Yang, Danai Koutra
摘要
Graph classification has a wide range of applications in bioinformatics, social sciences, automated fake news detection, web document classification, and more. In many practical scenarios, including web-scale applications, labels are scarce or hard to obtain. Unsupervised learning is thus a natural paradigm for these settings, but its performance often lags behind that of supervised learning. However, recently contrastive learning (CL) has enabled unsupervised computer vision models to perform comparably to supervised models. Theoretical and empirical works analyzing visual CL frameworks find that leveraging large datasets and task relevant augmentations is essential for CL framework success. Interestingly, graph CL frameworks report high performance while using orders of magnitude smaller data, and employing domain-agnostic graph augmentations (DAGAs) that can corrupt task relevant information. Motivated by these discrepancies, we seek to determine why existing graph CL frameworks continue to perform well, and identify flawed practices in graph data augmentation and popular graph CL evaluation protocols. We find that DAGA can destroy task-relevant information and harm the model’s ability to learn discriminative representations. We also show that on small benchmark datasets, the inductive bias of graph neural networks can significantly compensate for these limitations, while on larger graph classification tasks commonly-used DAGAs perform poorly. Based on our findings, we propose better practices and sanity checks for future research and applications, including adhering to principles in visual CL when designing context-aware graph augmentations. For example, in graph-based document classification, which can be used for better web search, we show task-relevant augmentations improve accuracy by up to 20.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- GraphMAE2: A Decoding-Enhanced Masked Self-Supervised Graph LearnerZhenyu Hou, Yufei He, Yukuo Cen, Xiao Liu 等WWW 2023 · 被引用 183 次
- Augmentations in Hypergraph Contrastive Learning: Fabricated and GenerativeTianxin Wei, Yuning You, Tianlong Chen, Yang Shen 等NeurIPS 2022 · 被引用 96 次
- Decoupled Self-supervised Learning for GraphsTeng Xiao, Zhengyu Chen, Zhimeng Guo, Zeyang Zhuang 等NeurIPS 2022 · 被引用 75 次
- Automated Spatio-Temporal Graph Contrastive LearningQianru Zhang, Chao Huang, Lianghao Xia, Zheng Wang 等WWW 2023 · 被引用 72 次
- Mechanistic Mode ConnectivityEkdeep Singh Lubana, Eric J. Bigelow, Robert P. Dick, David Scott Krueger 等ICML 2023 · 被引用 57 次
它引用的顶会 Paper39
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec 等NeurIPS 2020 · 被引用 9,171 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
- Graph Contrastive Learning with AugmentationsYuning You, Tianlong Chen, Yongduo Sui, Ting Chen 等NeurIPS 2020 · 被引用 3,042 次
相关 Paper
- Analyzing Data-Centric Properties for Graph Contrastive LearningPuja Trivedi, Ekdeep Singh Lubana, Mark Heimann, Danai Koutra 等NeurIPS 2022 · 被引用 13 次
- Adversarial Graph Contrastive Learning with Information RegularizationShengyu Feng, Baoyu Jing, Yada Zhu, Hanghang TongWWW 2022 · 被引用 76 次
- Architecture Matters: Uncovering Implicit Mechanisms in Graph Contrastive LearningXiaojun Guo, Yifei Wang, Zeming Wei, Yisen WangNeurIPS 2023 · 被引用 20 次
- Graph Contrastive Learning with Adaptive AugmentationYanqiao Zhu, Yichen Xu, Feng Yu, Qiang Liu 等WWW 2021 · 被引用 1,415 次
- Label-invariant Augmentation for Semi-Supervised Graph ClassificationHan Yue, Chunhui Zhang, Chuxu Zhang, Hongfu LiuNeurIPS 2022 · 被引用 37 次
