Conformal Prediction with Corrupted Labels: Uncertain Imputation and Robust Re-weighting
Shai Feldman, Stephen Bates, Yaniv Romano
Abstract
We introduce a framework for robust uncertainty quantification in situations where labeled training data are corrupted, through noisy or missing labels. We build on conformal prediction, a statistical tool for generating prediction sets that cover the test label with a pre-specified probability. The validity of conformal prediction, however, holds under the i.i.d assumption, which does not hold in our setting due to the corruptions in the data. To account for this distribution shift, the privileged conformal prediction (PCP) method proposed leveraging privileged information (PI) -- additional features available only during training -- to re-weight the data distribution, yielding valid prediction sets under the assumption that the weights are accurate. In this work, we analyze the robustness of PCP to inaccuracies in the weights. Our analysis indicates that PCP can still yield valid uncertainty estimates even when the weights are poorly estimated. Furthermore, we introduce uncertain imputation (UI), a new conformal method that does not rely on weight estimation. Instead, we impute corrupted labels in a way that preserves their uncertainty. Our approach is supported by theoretical guarantees and validated empirically on both synthetic and real benchmarks. Finally, we show that these techniques can be integrated into a triply robust framework, ensuring statistically valid predictions as long as at least one underlying method is valid.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on11
- Adaptive Conformal Inference Under Distribution ShiftIsaac Gibbs, Emmanuel J. CandèsNeurIPS 2021 · 665 citations
- Adaptive Conformal Predictions for Time SeriesMargaux Zaffran, Olivier Féron, Yannig Goude, Julie Josse et al.ICML 2022 · 209 citations
- Deep Proxy Causal Learning and its Application to Confounded Bandit Policy EvaluationLiyuan Xu, Heishiro Kanagawa, Arthur GrettonNeurIPS 2021 · 52 citations
- Toward Understanding Privileged Features Distillation in Learning-to-RankShuo Yang, Sujay Sanghavi, Holakou Rahmanian, Jan Bakus et al.NeurIPS 2022 · 31 citations
- Conformal Prediction with Missing ValuesMargaux Zaffran, Aymeric Dieuleveut, Julie Josse, Yaniv RomanoICML 2023 · 31 citations
Related papers
- Robust Conformal Prediction Using Privileged InformationShai Feldman, Yaniv RomanoNeurIPS 2024 · 7 citations
- Conformal Prediction for Class-wise Coverage via Augmented Label Rank CalibrationYuanjie Shi, Subhankar Ghosh, Taha Belkhouja, Jana Doppa et al.NeurIPS 2024 · 29 citations
- Conformal Prediction with Cellwise Outliers: A Detect-then-Impute ApproachQian Peng, Yajie Bao, Haojie Ren, Zhaojun Wang et al.ICML 2025
- Causality-Based Conformal Imputation Correction with Non-Random Missing LabelsChunyuan Zheng, Xiang Li, Hang Pan, Eric Wang et al.KDD 2026
- Conformal Prediction for Federated Uncertainty Quantification Under Label ShiftVincent Plassier, Mehdi Makni, Aleksandr Rubashevskii, Eric Moulines et al.ICML 2023 · 28 citations
