ValUES: A Framework for Systematic Validation of Uncertainty Estimation in Semantic Segmentation
Kim-Celine Kahl, Carsten T. Lüth, Maximilian Zenk, Klaus H. Maier-Hein, Paul F. Jaeger
摘要
Uncertainty estimation is an essential and heavily-studied component for the reliable application of semantic segmentation methods. While various studies exist claiming methodological advances on the one hand, and successful application on the other hand, the field is currently hampered by a gap between theory and practice leaving fundamental questions unanswered: Can data-related and model-related uncertainty really be separated in practice? Which components of an uncertainty method are essential for real-world performance? Which uncertainty method works well for which application? In this work, we link this research gap to a lack of systematic and comprehensive evaluation of uncertainty methods. Specifically, we identify three key pitfalls in current literature and present an evaluation framework that bridges the research gap by providing 1) a controlled environment for studying data ambiguities as well as distribution shifts, 2) systematic ablations of relevant method components, and 3) test-beds for the five predominant uncertainty applications: OoD-detection, active learning, failure detection, calibration, and ambiguity modeling. Empirical results on simulated as well as real-world data demonstrate how the proposed framework is able to answer the predominant questions in the field revealing for instance that 1) separation of uncertainty types works on simulated data but does not necessarily translate to real-world data, 2) aggregation of scores is a crucial but currently neglected component of uncertainty methods, 3) While ensembles are performing most robustly across the different downstream tasks and settings, test-time augmentation often constitutes a light-weight alternative. Code is at: https://github.com/IML-DKFZ/values
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Better than Average: Spatially-Aware Aggregation of Segmentation Uncertainty Improves Downstream PerformanceVanessa Emanuela Guarino, Claudia Winklmayr, Jannik Franzen, Josef Rumberger 等CVPR 2026 · 被引用 1 次
- REPEAT: Improving Uncertainty Estimation in Representation Learning ExplainabilityKristoffer K. Wickstrøm, Thea Brüsch, Michael C. Kampffmeyer, Robert JenssenAAAI 2025
它引用的顶会 Paper6
- Sampling-Free Epistemic Uncertainty Estimation Using Approximated Variance PropagationJanis Postels, Francesco Ferroni, Huseyin Coskun, Nassir Navab 等ICCV 2019 · 被引用 153 次
- Stochastic Segmentation Networks: Modelling Spatially Correlated Aleatoric UncertaintyMiguel Monteiro, Loïc Le Folgoc, Daniel Coelho de Castro, Nick Pawlowski 等NeurIPS 2020 · 被引用 153 次
- On the Practicality of Deterministic Epistemic UncertaintyJanis Postels, Mattia Segù, Tao Sun, Luca Daniel Sieber 等ICML 2022 · 被引用 76 次
- Navigating the Pitfalls of Active Learning Evaluation: A Systematic Framework for Meaningful Performance AssessmentCarsten T. Lüth, Till J. Bungert, Lukas Klein, Paul F. JaegerNeurIPS 2023 · 被引用 32 次
- A Call to Reflect on Evaluation Practices for Failure Detection in Image ClassificationPaul F. Jaeger, Carsten T. Lüth, Lukas Klein, Till J. BungertICLR 2023 · 被引用 18 次
相关 Paper
- Adaptive Dual Uncertainty Optimization: Boosting Monocular 3D Object Detection under Test-Time ShiftsZixuan Hu, Dongxiao Li, Xinzhu Ma, Shixiang Tang 等ICCV 2025
- Generalize or Detect? Towards Robust Semantic Segmentation Under Multiple Distribution ShiftsZhitong Gao, Bingnan Li, Mathieu Salzmann, Xuming HeNeurIPS 2024 · 被引用 10 次
- Modeling the Distributional Uncertainty for Salient Object Detection ModelsXinyu Tian, Jing Zhang, Mochu Xiang, Yuchao DaiCVPR 2023
- UncertainSAM: Fast and Efficient Uncertainty Quantification of the Segment Anything ModelTimo Kaiser, Thomas Norrenbrock, Bodo RosenhahnICML 2025
- Guided Curriculum Model Adaptation and Uncertainty-Aware Evaluation for Semantic Nighttime Image SegmentationChristos Sakaridis, Dengxin Dai, Luc Van GoolICCV 2019 · 被引用 297 次
