Reliable and Trustworthy Machine Learning for Health Using Dataset Shift Detection
Chunjong Park, Anas Awadalla, Tadayoshi Kohno, Shwetak N. Patel
摘要
Unpredictable ML model behavior on unseen data, especially in the health domain, raises serious concerns about its safety as repercussions for mistakes can be fatal. In this paper, we explore the feasibility of using state-of-the-art out-of-distribution detectors for reliable and trustworthy diagnostic predictions. We select publicly available deep learning models relating to various health conditions (e.g., skin cancer, lung sound, and Parkinson's disease) using various input data types (e.g., image, audio, and motion data). We demonstrate that these models show unreasonable predictions on out-of-distribution datasets. We show that Mahalanobis distance- and Gram matrices-based out-of-distribution detection methods are able to detect out-of-distribution data with high accuracy for the health models that operate on different modalities. We then translate the out-of-distribution score into a human interpretable CONFIDENCE SCORE to investigate its effect on the users' interaction with health ML applications. Our user study shows that the confidence score helped the participants only trust the results with a high score to make a medical decision and disregard results with a low score. Through this work, we demonstrate that dataset shift is a critical piece of information for high-stake ML applications, such as medical diagnosis and healthcare, to provide reliable and trustworthy predictions to the users.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Investigating Generalizability of Speech-based Suicidal Ideation Detection Using Mobile PhonesArvind Pillai, Subigya Kumar Nepal, Weichen Wang, Matthew Nemesure 等UbiComp 2024 · 被引用 26 次
- Data-SUITE: Data-centric identification of in-distribution incongruous examplesNabeel Seedat, Jonathan Crabbé, Mihaela van der SchaarICML 2022 · 被引用 16 次
- ProfiliTable: Profiling-Driven Tabular Data Processing via Agentic WorkflowsWei Liu, Yang Gu, Xi Yan, Zihan Nan 等KDD 2026 · 被引用 1 次
- The Best of Both Worlds: On the Dilemma of Out-of-distribution DetectionQingyang Zhang, Qiuxuan Feng, Joey Tianyi Zhou, Yatao Bian 等NeurIPS 2024
它引用的顶会 Paper8
- Deep Learning with Differential PrivacyMartín Abadi, Andy Chu, Ian J. Goodfellow, H. Brendan McMahan 等CCS 2016 · 被引用 7,620 次
- Energy-based Out-of-distribution DetectionWeitang Liu, Xiaoyun Wang, John D. Owens, Yixuan LiNeurIPS 2020 · 被引用 2,213 次
- Exploiting Unintended Feature Leakage in Collaborative LearningLuca Melis, Congzheng Song, Emiliano De Cristofaro, Vitaly ShmatikovS&P 2019 · 被引用 1,736 次
- Interpreting Interpretability: Understanding Data Scientists' Use of Interpretability Tools for Machine LearningHarmanpreet Kaur, Harsha Nori, Samuel Jenkins, Rich Caruana 等CHI 2020 · 被引用 541 次
- Multi-Task Temporal Shift Attention Networks for On-Device Contactless Vitals MeasurementXin Liu, Josh Fromm, Shwetak N. Patel, Daniel McDuffNeurIPS 2020 · 被引用 436 次
相关 Paper
- Towards Trustable Skin Cancer Diagnosis via Rewriting Model's DecisionSiyuan Yan, Zhen Yu, Xuelin Zhang, Dwarikanath Mahapatra 等CVPR 2023
- OpenMIBOOD: Open Medical Imaging Benchmarks for Out-Of-Distribution DetectionMax Gutbrod, David Rauber, Danilo Weber Nunes, Christoph PalmCVPR 2025
- Towards a Certificate of Trust: Task-Aware OOD Detection for Scientific AIBogdan Raonic, Siddhartha Mishra, Samuel LanthalerICLR 2026 · 被引用 2 次
- DIsoN: Decentralized Isolation Networks for Out-of-Distribution Detection in Medical ImagingFelix Wagner, Pramit Saha, Harry Anthony, J. Alison Noble 等NeurIPS 2025 · 被引用 1 次
- The Invisible Gorilla Effect in Out-of-distribution DetectionHarry Anthony, Ziyun Liang, Hermione Warr, Konstantinos KamnitsasCVPR 2026
