IceBerg: Debiased Self-Training for Class-Imbalanced Node Classification
Zhixun Li, Dingshuo Chen, Tong Zhao, Daixin Wang, Hongrui Liu, Zhiqiang Zhang, Jun Zhou, Jeffrey Xu Yu
Abstract
Graph Neural Networks (GNNs) have achieved great success in dealing with non-Euclidean graph-structured data and have been widely deployed in many real-world applications. However, their effectiveness is often jeopardized under class-imbalanced training sets. Most existing studies have analyzed class-imbalanced node classification from a supervised learning perspective, they do not fully utilize the large number of unlabeled nodes in semi-supervised scenarios. We claim that the supervised signal is just the tip of the iceberg and a large number of unlabeled nodes have not yet been effectively utilized. In this work, we propose IceBerg, a debiased self-training framework to address the class-imbalanced and few-shot challenges for GNNs at the same time. Specifically, to figure out the Matthew effect and label distribution shift in self-training, we propose Double Balancing, which can largely improve the performance of existing baselines with just a few lines of code as a simple plug-and-play module. Secondly, to enhance the long-range propagation capability of GNNs, we disentangle the propagation and transformation operations of GNNs. Therefore, the weak supervision signals can propagate more effectively to address the few-shot issue. In summary, we find that leveraging unlabeled nodes can significantly enhance the performance of GNNs in class-imbalanced and few-shot scenarios, and even small, surgical modifications can lead to substantial performance improvements. Systematic experiments on benchmark datasets show that our method can deliver considerable performance gain over existing class-imbalanced node classification baselines. Additionally, due to IceBerg's outstanding ability to leverage unsupervised signals, it also achieves state-of-the-art results in few-shot node classification scenarios. The code of IceBerg is available at: https://github.com/ZhixunLEE/IceBerg.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a02484fc-9d94-45e9-a66f-63cfbe1e6519Cited by top-tier papers4
- Preference-driven Knowledge Distillation for Few-shot Node ClassificationXing Wei, Chunchun Chen, Rui Fan, Xiaofeng Cao et al.NeurIPS 2025 · 2 citations
- Geometric Imbalance in Semi-Supervised Node ClassificationLiang Yan, Shengzhong Zhang, Bisheng Li, Menglin Yang et al.NeurIPS 2025 · 2 citations
- IGL-Bench: Establishing the Comprehensive Benchmark for Imbalanced Graph LearningJiawen Qin, Haonan Yuan, Qingyun Sun, Lyujin Xu et al.ICLR 2025
- Rationalizing and Augmenting Dynamic Graph Neural NetworksGuibin Zhang, Yiyan Qi, Ziyang Cheng, Yanwei Yue et al.ICLR 2025
Builds on30
- Decoupling Representation and Classifier for Long-Tailed RecognitionBingyi Kang, Saining Xie, Marcus Rohrbach, Zhicheng Yan et al.ICLR 2020 · 1,496 citations
- Symmetric Cross Entropy for Robust Learning With Noisy LabelsYisen Wang, Xingjun Ma, Zaiyi Chen, Yuan Luo et al.ICCV 2019 · 1,125 citations
- Balanced Meta-Softmax for Long-Tailed Visual RecognitionJiawei Ren, Cunjun Yu, Shunan Sheng, Xiao Ma et al.NeurIPS 2020 · 861 citations
- Towards Deeper Graph Neural NetworksMeng Liu, Hongyang Gao, Shuiwang JiKDD 2020 · 496 citations
- Multi-Stage Self-Supervised Learning for Graph Convolutional Networks on Graphs with Few Labeled NodesKe Sun, Zhouchen Lin, Zhanxing ZhuAAAI 2020 · 304 citations
Related papers
- Meta Propagation Networks for Graph Few-shot Semi-supervised LearningKaize Ding, Jianling Wang, James Caverlee, Huan LiuAAAI 2022 · 56 citations
- BIM: Improving Graph Neural Networks with Balanced Influence MaximizationWentao Zhang, Xinyi Gao, Ling Yang, Meng Cao et al.ICDE 2024 · 4 citations
- Rethinking Semi-Supervised Imbalanced Node Classification from Bias-Variance DecompositionDivin Yan, Gengchen Wei, Chen Yang, Shengzhong Zhang et al.NeurIPS 2023 · 27 citations
- Edge Prompt Tuning for Graph Neural NetworksXingbo Fu, Yinhan He, Jundong LiICLR 2025 · 140 citations
- GPPT: Graph Pre-training and Prompt Tuning to Generalize Graph Neural NetworksMingchen Sun, Kaixiong Zhou, Xin He, Ying Wang et al.KDD 2022 · 141 citations
