Mitigating Label Noise on Graphs via Topological Sample Selection
Yuhao Wu, Jiangchao Yao, Xiaobo Xia, Jun Yu, Ruxin Wang, Bo Han, Tongliang Liu
Abstract
Despite the success of the carefully-annotated benchmarks, the effectiveness of existing graph neural networks (GNNs) can be considerably impaired in practice when the real-world graph data is noisily labeled. Previous explorations in sample selection have been demonstrated as an effective way for robust learning with noisy labels, however, the conventional studies focus on i.i.d data, and when moving to non-iid graph data and GNNs, two notable challenges remain: (1) nodes located near topological class boundaries are very informative for classification but cannot be successfully distinguished by the heuristic sample selection. (2) there is no available measure that considers the graph topological information to promote sample selection in a graph. To address this dilemma, we propose a Topological Sample Selection (TSS) method that boosts the informative sample selection process in a graph by utilising topological information. We theoretically prove that our procedure minimizes an upper bound of the expected risk under target clean distribution, and experimentally show the superiority of our method compared with state-ofthe-art baselines. Our implementation is available at https://github.com/tmllab/2024_ICML_TSS .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e0c21087-a557-4cb4-9da6-b38505e00dd4Cited by top-tier papers7
- Geometric Imbalance in Semi-Supervised Node ClassificationLiang Yan, Shengzhong Zhang, Bisheng Li, Menglin Yang et al.NeurIPS 2025 · 2 citations
- Training Robust Graph Neural Networks by Modeling Noise DependenciesYeonjun In, Kanghoon Yoon, Sukwon Yun, Kibum Kim et al.NeurIPS 2025 · 2 citations
- AutoGFM: Automated Graph Foundation Model with Adaptive Architecture CustomizationHaibo Chen, Xin Wang, Zeyang Zhang, Haoyang Li et al.ICML 2025
- Instance-dependent Early StoppingSuqin Yuan, Runqi Lin, Lei Feng, Bo Han et al.ICLR 2025
- Towards Out-of-Modal Generalization without Instance-level Modal CorrespondenceZhuo Huang, Gang Niu, Bo Han, Masashi Sugiyama et al.ICLR 2025
Builds on41
- Graph Contrastive Learning with AugmentationsYuning You, Tianlong Chen, Yongduo Sui, Ting Chen et al.NeurIPS 2020 · 3,042 citations
- Beyond Homophily in Graph Neural Networks: Current Limitations and Effective DesignsJiong Zhu, Yujun Yan, Lingxiao Zhao, Mark Heimann et al.NeurIPS 2020 · 1,490 citations
- MAGNN: Metapath Aggregated Graph Neural Network for Heterogeneous Graph EmbeddingXinyu Fu, Jiani Zhang, Ziqiao Meng, Irwin KingWWW 2020 · 1,149 citations
- Graph Information BottleneckTailin Wu, Hongyu Ren, Pan Li, Jure LeskovecNeurIPS 2020 · 366 citations
- SELF: Learning to Filter Noisy Labels with Self-EnsemblingDuc Tam Nguyen, Chaithanya Kumar Mummadi, Thi-Phuong-Nhung Ngo, Thi Hoai Phuong Nguyen et al.ICLR 2020 · 354 citations
Related papers
- NRGNN: Learning a Label Noise Resistant Graph Neural Network on Sparsely and Noisily Labeled GraphsEnyan Dai, Charu Aggarwal, Suhang WangKDD 2021 · 80 citations
- Identifying and Correcting Label Noise for Robust GNNs via Influence ContradictionWei Ju, Wei Zhang, Siyu Yi, Zhengyang Mao et al.ICML 2026
- Shift-Robust GNNs: Overcoming the Limitations of Localized Graph Training dataQi Zhu, Natalia Ponomareva, Jiawei Han, Bryan PerozziNeurIPS 2021 · 152 citations
- Can Pseudo-Label Be More Reliable? A Simple yet Effective Topology-Aware Graph Self-Training MethodGen Liu, Zhongying Zhao, Hui Zhou, Chao Li et al.AAAI 2026
- Let Your Features Tell The Differences: Understanding Graph Convolution By Feature SplittingYilun Zheng, Xiang Li, Sitao Luan, Xiaojiang Peng et al.ICLR 2025
