When Sparsity Meets Contrastive Models: Less Graph Data Can Bring Better Class-Balanced Representations
Chunhui Zhang, Chao Huang, Yijun Tian, Qianlong Wen, Zhongyu Ouyang, Youhuan Li, Yanfang Ye, Chuxu Zhang
Abstract
Graph Neural Networks (GNNs) are powerful models for non-Euclidean data, but their training is often accentuated by massive unnecessary computation: On the one hand, training on non-Euclidean data has relatively high computational cost due to its irregular density properties; on the other hand, the class imbalance property often associated with non-Euclidean data cannot be alleviated by the massiveness of the data, thus hindering the generalisation of the models. To address the above issues, theoretically, we start with a hypothesis about the effectiveness of using a subset of training data for GNNs, which is guaranteed by the gradient distance between the subset and the full set. Empirically, we also observe that a subset of the data can provide informative gradients for model optimization and which changes over time dynamically. We name this phenomenon dynamic data sparsity. Additionally, we find that pruned sparse contrastive models may "miss" valuable information, leading to a large loss value on the informative subset. Motivated by the above findings, we develop a unified data model dynamic sparsity framework called Data Decantation (DataDec) to address the above challenges. The key idea of DataDec is to identify the informative subset dynamically during the training process by applying sparse graph contrastive learning. The effectiveness of DataDec is comprehensively evaluated on graph benchmark datasets and we also verify its generalizability on image data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext aa17d304-01b2-46c4-b319-20017061d038Cited by top-tier papers9
- Aligning Relational Learning with Lipschitz FairnessYaning Jia, Chunhui Zhang, Soroush VosoughiICLR 2024 · 11 citations
- Mitigating Emergent Robustness Degradation while Scaling Graph LearningXiangchi Yuan, Chunhui Zhang, Yijun Tian, Yanfang Ye et al.ICLR 2024 · 10 citations
- Mastering Long-Tail Complexity on Graphs: Characterization, Learning, and GeneralizationHaohui Wang, Baoyu Jing, Kaize Ding, Yada Zhu et al.KDD 2024 · 7 citations
- Multi-graph Fusion Cross-model Contrastive Learning for RecommendationShengjun Ma, Yuhai Zhao, Fenglong Ma, Baoyin Liu et al.AAAI 2026
- SamGoG: A Sampling-Based Graph-of-Graphs Framework for Imbalanced Graph ClassificationShangyou Wang, Zezhong Ding, Xike XieKDD 2026
Builds on19
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong et al.NeurIPS 2020 · 3,935 citations
- Graph Contrastive Learning with AugmentationsYuning You, Tianlong Chen, Yongduo Sui, Ting Chen et al.NeurIPS 2020 · 3,042 citations
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan et al.ICLR 2020 · 1,496 citations
- InfoGraph: Unsupervised and Semi-supervised Graph-Level Representation Learning via Mutual Information MaximizationFan-Yun Sun, Jordan Hoffmann, Vikas Verma, Jian TangICLR 2020 · 1,010 citations
- Pruning neural networks without any data by iteratively conserving synaptic flowHidenori Tanaka, Daniel Kunin, Daniel L. K. Yamins, Surya GanguliNeurIPS 2020 · 884 citations
Related papers
- Uncovering Capabilities of Model Pruning in Graph Contrastive LearningJunran Wu, Xueyuan Chen, Shangzhe LiACM MM 2024 · 2 citations
- Dual-Kernel Graph Community Contrastive LearningXiang Chen, Kun Yue, Wenjie Liu, Zhenyu Zhang et al.AAAI 2026
- Two Heads Are Better Than One: Boosting Graph Sparse Training via Semantic and Topological AwarenessGuibin Zhang, Yanwei Yue, Kun Wang, Junfeng Fang et al.ICML 2024 · 15 citations
- GDeR: Safeguarding Efficiency, Balancing, and Robustness via Prototypical Graph PruningGuibin Zhang, Haonan Dong, Yuchen Zhang, Zhixun Li et al.NeurIPS 2024 · 9 citations
- StructComp: Substituting propagation with Structural Compression in Training Graph Contrastive LearningShengzhong Zhang, Wenjie Yang, Xinyuan Cao, Hongwei Zhang et al.ICLR 2024 · 6 citations
