Towards a Certificate of Trust: Task-Aware OOD Detection for Scientific AI
Bogdan Raonic, Siddhartha Mishra, Samuel Lanthaler
摘要
Data-driven models are increasingly adopted in critical scientific fields like weather forecasting and fluid dynamics. These methods can fail on out-of-distribution (OOD) data, but detecting such failures in regression tasks is an open challenge. We propose a new OOD detection method based on estimating joint likelihoods using a score-based diffusion model. This approach considers not just the input but also the regression model's prediction, providing a task-aware reliability score. Across numerous scientific datasets, including PDE datasets, satellite imagery and brain tumor segmentation, we show that this likelihood strongly correlates with prediction error. Our work provides a foundational step towards building a verifiable 'certificate of trust', thereby offering a practical tool for assessing the trustworthiness of AI-based scientific predictions. Our code is publicly available at https://github.com/bogdanraonic3/OOD_Detection_ScientificML
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper18
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Elucidating the Design Space of Diffusion-Based Generative ModelsTero Karras, Miika Aittala, Timo Aila, Samuli LaineNeurIPS 2022 · 被引用 3,959 次
- Fourier Neural Operator for Parametric Partial Differential EquationsZongyi Li, Nikola Borislavov Kovachki, Kamyar Azizzadenesheli, Burigede Liu 等ICLR 2021 · 被引用 3,911 次
- Energy-based Out-of-distribution DetectionWeitang Liu, Xiaoyun Wang, John D. Owens, Yixuan LiNeurIPS 2020 · 被引用 2,213 次
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar 等ICLR 2021 · 被引用 1,270 次
相关 Paper
- KLIP: Localized Distribution Shift Detection via KL-Divergence with Diffusion Priors in Inverse ProblemsAlireza Kheirandish, Jihoon Hong, Sara Fridovich-KeilCVPR 2026
- Beyond the Norms: Detecting Prediction Errors in Regression ModelsAndrés Altieri, Marco Romanelli, Georg Pichler, Florence Alberge 等ICML 2024 · 被引用 1 次
- Reliable and Trustworthy Machine Learning for Health Using Dataset Shift DetectionChunjong Park, Anas Awadalla, Tadayoshi Kohno, Shwetak N. PatelNeurIPS 2021 · 被引用 50 次
- A Geometric Explanation of the Likelihood OOD Detection ParadoxHamidreza Kamkari, Brendan Leigh Ross, Jesse C. Cresswell, Anthony L. Caterini 等ICML 2024 · 被引用 20 次
- EigenScore: OOD Detection using Posterior Covariance in Diffusion ModelsShirin Shoushtari, Yi Wang, Xiao Shi, M. Salman Asif 等ICLR 2026 · 被引用 5 次
