Towards Reliable Rare Category Analysis on Graphs via Individual Calibration
Longfeng Wu, Bowen Lei, Dongkuan Xu, Dawei Zhou
Abstract
Rare categories abound in a number of real-world networks and play a pivotal role in a variety of high-stakes applications, including financial fraud detection, network intrusion detection, and rare disease diagnosis. Rare category analysis (RCA) refers to the task of detecting, characterizing, and comprehending the behaviors of minority classes in a highly-imbalanced data distribution. While the vast majority of existing work on RCA has focused on improving the prediction performance, a few fundamental research questions heretofore have received little attention and are less explored: How confident or uncertain is a prediction model in rare category analysis? How can we quantify the uncertainty in the learning process and enable reliable rare category analysis? To answer these questions, we start by investigating miscalibration in existing RCA methods. Empirical results reveal that stateof-the-art RCA methods are mainly over-confident in predicting minority classes and under-confident in predicting majority classes. Motivated by the observation, we propose a novel individual calibration framework, named CaliRare, for alleviating the unique challenges of RCA, thus enabling reliable rare category analysis. In particular, to quantify the uncertainties in RCA, we develop a node-level uncertainty quantification algorithm to model the overlapping support regions with high uncertainty; to handle the rarity of minority classes in miscalibration calculation, we generalize the distribution-based calibration metric to the instance level and propose the first individual calibration measurement on graphs named Expected Individual Calibration Error (EICE). We perform extensive experimental evaluations on real-world datasets, including rare category characterization and model calibration tasks, which demonstrate the significance of our proposed framework.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext af509799-9ca7-45ca-b9bd-b2742d2e4cf6Cited by top-tier papers2
- Rethinking Data Distillation: Do Not Overlook CalibrationDongyao Zhu, Yanbo Fang, Bowen Lei, Yiqun Xie et al.ICCV 2023 · 19 citations
- Bridging Fairness and Uncertainty: Theoretical Insights and Practical Strategies for Equalized Coverage in GNNsLongfeng Wu, Yao Zhou, Jian Kang, Dawei ZhouWWW 2025 · 4 citations
Builds on10
- MAGNN: Metapath Aggregated Graph Neural Network for Heterogeneous Graph EmbeddingXinyu Fu, Jiani Zhang, Ziqiao Meng, Irwin KingWWW 2020 · 1,149 citations
- Revisiting the Calibration of Modern Neural NetworksMatthias Minderer, Josip Djolonga, Rob Romijnders, Frances Hubis et al.NeurIPS 2021 · 633 citations
- Pitfalls of In-Domain Uncertainty Estimation and Ensembling in Deep LearningArsenii Ashukha, Alexander Lyzhov, Dmitry Molchanov, Dmitry P. VetrovICLR 2020 · 354 citations
- Mix-n-Match : Ensemble and Compositional Methods for Uncertainty Calibration in Deep LearningJize Zhang, Bhavya Kailkhura, Thomas Yong-Jin HanICML 2020 · 276 citations
- Mixup for Node and Graph ClassificationYiwei Wang, Wei Wang, Yuxuan Liang, Yujun Cai et al.WWW 2021 · 220 citations
Related papers
- ALLIE: Active Learning on Large-scale Imbalanced GraphsLimeng Cui, Xianfeng Tang, Sumeet Katariya, Nikhil Rao et al.WWW 2022 · 25 citations
- Calibrate: Interactive Analysis of Probabilistic Model OutputPeter Xenopoulos, João Rulff, Luis Gustavo Nonato, Brian Barr et al.IEEE VIS 2022 · 19 citations
- Fractal Calibration for Long-tailed Object DetectionKonstantinos Panagiotis Alexandridis, Ismail Elezi, Jiankang Deng, Anh Nguyen et al.CVPR 2025
- GIER: Addressing Class Imbalance in GNNs Through Experience ReplayLiu Yang, Chuyao Liu, Zidong Wang, Tingxuan Chen et al.AAAI 2026
- Variable-Based Calibration for Machine Learning ClassifiersMarkelle Kelly, Padhraic SmythAAAI 2023 · 7 citations
