Robust Calibration with Multi-domain Temperature Scaling
Yaodong Yu, Stephen Bates, Yi Ma, Michael I. Jordan
摘要
Uncertainty quantification is essential for the reliable deployment of machine learning models to high-stakes application domains. Uncertainty quantification is all the more challenging when training distribution and test distribution are different, even the distribution shifts are mild. Despite the ubiquity of distribution shifts in real-world applications, existing uncertainty quantification approaches mainly study the in-distribution setting where the train and test distributions are the same. In this paper, we develop a systematic calibration model to handle distribution shifts by leveraging data from multiple domains. Our proposed method -- multi-domain temperature scaling -- uses the heterogeneity in the domains to improve calibration robustness under distribution shift. Through experiments on three benchmark data sets, we find our proposed method outperforms existing methods as measured on both in-distribution and out-of-distribution test sets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper20
- Towards Calibrated Robust Fine-Tuning of Vision-Language ModelsChangdae Oh, Hyesu Lim, Mijoo Kim, Dongyoon Han 等NeurIPS 2024 · 被引用 49 次
- Federated Conformal Predictors for Distributed Uncertainty QuantificationCharles Lu, Yaodong Yu, Sai Praneeth Karimireddy, Michael I. Jordan 等ICML 2023 · 被引用 47 次
- Thermometer: Towards Universal Calibration for Large Language ModelsMaohao Shen, Subhro Das, Kristjan H. Greenewald, Prasanna Sattigeri 等ICML 2024 · 被引用 38 次
- An Empirical Study Into What Matters for Calibrating Vision-Language ModelsWeijie Tu, Weijian Deng, Dylan Campbell, Stephen Gould 等ICML 2024 · 被引用 18 次
- Open-Vocabulary Calibration for Fine-tuned CLIPShuoyuan Wang, Jindong Wang, Guoqing Wang, Bob Zhang 等ICML 2024 · 被引用 17 次
它引用的顶会 Paper11
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- The Many Faces of Robustness: A Critical Analysis of Out-of-Distribution GeneralizationDan Hendrycks, Steven Basart, Norman Mu, Saurav Kadavath 等ICCV 2021 · 被引用 2,294 次
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie 等ICML 2021 · 被引用 1,773 次
- AugMix: A Simple Data Processing Method to Improve Robustness and UncertaintyDan Hendrycks, Norman Mu, Ekin Dogus Cubuk, Barret Zoph 等ICLR 2020 · 被引用 1,572 次
- Revisiting the Calibration of Modern Neural NetworksMatthias Minderer, Josip Djolonga, Rob Romijnders, Frances Hubis 等NeurIPS 2021 · 被引用 633 次
相关 Paper
- Consistency-Guided Temperature Scaling Using Style and Content Information for Out-of-Domain CalibrationWonjeong Choi, Jungwuk Park, Dong-Jun Han, Younghyun Park 等AAAI 2024 · 被引用 3 次
- Post-Hoc Uncertainty Calibration for Domain Drift ScenariosChristian Tomani, Sebastian Gruber, Muhammed Ebrar Erdem, Daniel Cremers 等CVPR 2021
- Confidence Calibration for Domain Generalization under Covariate ShiftYunye Gong, Xiao Lin, Yi Yao, Thomas G. Dietterich 等ICCV 2021 · 被引用 35 次
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang 等ICML 2021 · 被引用 1,163 次
- Robustness via Cross-Domain EnsemblesTeresa Yeo, Oguzhan Fatih Kar, Amir ZamirICCV 2021 · 被引用 30 次
