When Sparsity Meets Contrastive Models: Less Graph Data Can Bring Better Class-Balanced Representations
Chunhui Zhang, Chao Huang, Yijun Tian, Qianlong Wen, Zhongyu Ouyang, Youhuan Li, Yanfang Ye, Chuxu Zhang
摘要
Graph Neural Networks (GNNs) are powerful models for non-Euclidean data, but their training is often accentuated by massive unnecessary computation: On the one hand, training on non-Euclidean data has relatively high computational cost due to its irregular density properties; on the other hand, the class imbalance property often associated with non-Euclidean data cannot be alleviated by the massiveness of the data, thus hindering the generalisation of the models. To address the above issues, theoretically, we start with a hypothesis about the effectiveness of using a subset of training data for GNNs, which is guaranteed by the gradient distance between the subset and the full set. Empirically, we also observe that a subset of the data can provide informative gradients for model optimization and which changes over time dynamically. We name this phenomenon dynamic data sparsity. Additionally, we find that pruned sparse contrastive models may "miss" valuable information, leading to a large loss value on the informative subset. Motivated by the above findings, we develop a unified data model dynamic sparsity framework called Data Decantation (DataDec) to address the above challenges. The key idea of DataDec is to identify the informative subset dynamically during the training process by applying sparse graph contrastive learning. The effectiveness of DataDec is comprehensively evaluated on graph benchmark datasets and we also verify its generalizability on image data.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Aligning Relational Learning with Lipschitz FairnessYaning Jia, Chunhui Zhang, Soroush VosoughiICLR 2024 · 被引用 11 次
- Mitigating Emergent Robustness Degradation while Scaling Graph LearningXiangchi Yuan, Chunhui Zhang, Yijun Tian, Yanfang Ye 等ICLR 2024 · 被引用 10 次
- Mastering Long-Tail Complexity on Graphs: Characterization, Learning, and GeneralizationHaohui Wang, Baoyu Jing, Kaize Ding, Yada Zhu 等KDD 2024 · 被引用 7 次
- Multi-graph Fusion Cross-model Contrastive Learning for RecommendationShengjun Ma, Yuhai Zhao, Fenglong Ma, Baoyin Liu 等AAAI 2026
- SamGoG: A Sampling-Based Graph-of-Graphs Framework for Imbalanced Graph ClassificationShangyou Wang, Zezhong Ding, Xike XieKDD 2026
它引用的顶会 Paper19
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong 等NeurIPS 2020 · 被引用 3,935 次
- Graph Contrastive Learning with AugmentationsYuning You, Tianlong Chen, Yongduo Sui, Ting Chen 等NeurIPS 2020 · 被引用 3,042 次
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan 等ICLR 2020 · 被引用 1,496 次
- InfoGraph: Unsupervised and Semi-supervised Graph-Level Representation Learning via Mutual Information MaximizationFan-Yun Sun, Jordan Hoffmann, Vikas Verma, Jian TangICLR 2020 · 被引用 1,010 次
- Pruning neural networks without any data by iteratively conserving synaptic flowHidenori Tanaka, Daniel Kunin, Daniel L. K. Yamins, Surya GanguliNeurIPS 2020 · 被引用 884 次
相关 Paper
- Uncovering Capabilities of Model Pruning in Graph Contrastive LearningJunran Wu, Xueyuan Chen, Shangzhe LiACM MM 2024 · 被引用 2 次
- Dual-Kernel Graph Community Contrastive LearningXiang Chen, Kun Yue, Wenjie Liu, Zhenyu Zhang 等AAAI 2026
- Two Heads Are Better Than One: Boosting Graph Sparse Training via Semantic and Topological AwarenessGuibin Zhang, Yanwei Yue, Kun Wang, Junfeng Fang 等ICML 2024 · 被引用 15 次
- GDeR: Safeguarding Efficiency, Balancing, and Robustness via Prototypical Graph PruningGuibin Zhang, Haonan Dong, Yuchen Zhang, Zhixun Li 等NeurIPS 2024 · 被引用 9 次
- StructComp: Substituting propagation with Structural Compression in Training Graph Contrastive LearningShengzhong Zhang, Wenjie Yang, Xinyuan Cao, Hongwei Zhang 等ICLR 2024 · 被引用 6 次
