Consistency-Guided Temperature Scaling Using Style and Content Information for Out-of-Domain Calibration
Wonjeong Choi, Jungwuk Park, Dong-Jun Han, Younghyun Park, Jaekyun Moon
Abstract
Research interests in the robustness of deep neural networks against domain shifts have been rapidly increasing in recent years. Most existing works, however, focus on improving the accuracy of the model, not the calibration performance which is another important requirement for trustworthy AI systems. Temperature scaling (TS), an accuracy-preserving post-hoc calibration method, has been proven to be effective in in-domain settings, but not in out-of-domain (OOD) due to the difficulty in obtaining a validation set for the unseen domain beforehand. In this paper, we propose consistency-guided temperature scaling (CTS), a new temperature scaling strategy that can significantly enhance the OOD calibration performance by providing mutual supervision among data samples in the source domains. Motivated by our observation that over-confidence stemming from inconsistent sample predictions is the main obstacle to OOD calibration, we propose to guide the scaling process by taking consistencies into account in terms of two different aspects - style and content - which are the key components that can well-represent data samples in multi-domain settings. Experimental results demonstrate that our proposed strategy outperforms existing works, achieving superior OOD calibration performance on various datasets. This can be accomplished by employing only the source domains without compromising accuracy, making our scheme directly applicable to various trustworthy AI systems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Target-Agnostic Calibration under Distribution Shift with Frequency-Aware Gradient RectificationYilin Zhang, Cai Xu, You Wu, Ziyu Guan et al.ICML 2026 · 1 citation
- T-CIL: Temperature Scaling using Adversarial Perturbation for Calibration in Class-Incremental LearningSeonghyeon Hwang, Minsu Kim, Steven Euijong WhangCVPR 2025
- Temperature Scaling in Discrete Sequence (Language) ModelsHannah Scheufele, Peter Blohm, Vikas GargICML 2026
Builds on7
- Improving model calibration with accuracy versus uncertainty optimizationRanganath Krishnan, Omesh TickooNeurIPS 2020 · 217 citations
- The Devil is in the Margin: Margin-based Label Smoothing for Network CalibrationBingyuan Liu, Ismail Ben Ayed, Adrian Galdran, Jose DolzCVPR 2022 · 61 citations
- Robust Calibration with Multi-domain Temperature ScalingYaodong Yu, Stephen Bates, Yi Ma, Michael I. JordanNeurIPS 2022 · 58 citations
- Towards Trustworthy Predictions from Deep Neural Networks with Fast Adversarial CalibrationChristian Tomani, Florian BuettnerAAAI 2021 · 43 citations
- A Stitch in Time Saves Nine: A Train-Time Regularizing Loss for Improved Neural Network CalibrationRamya Hebbalaguppe, Jatin Prakash, Neelabh Madan, Chetan AroraCVPR 2022 · 38 citations
Related papers
- Post-Hoc Uncertainty Calibration for Domain Drift ScenariosChristian Tomani, Sebastian Gruber, Muhammed Ebrar Erdem, Daniel Cremers et al.CVPR 2021
- Calibrating Zero-shot Cross-lingual (Un-)structured PredictionsZhengping Jiang, Anqi Liu, Benjamin Van DurmeEMNLP 2022 · 4 citations
- Intra Order-preserving Functions for Calibration of Multi-Class Neural NetworksAmir Rahimi, Amirreza Shaban, Ching-An Cheng, Richard Hartley et al.NeurIPS 2020 · 96 citations
- Rethinking Calibration of Deep Neural Networks: Do Not Be Afraid of OverconfidenceDeng-Bao Wang, Lei Feng, Min-Ling ZhangNeurIPS 2021 · 177 citations
- Your Pre-trained LLM is Secretly an Unsupervised Confidence CalibratorBeier Luo, Shuoyuan Wang, Sharon Li, Hongxin WeiNeurIPS 2025 · 22 citations
