Controlling Counterfactual Harm in Decision Support Systems Based on Prediction Sets
Eleni Straitouri, Suhas Thejaswi, Manuel Gomez Rodriguez
摘要
Decision support systems based on prediction sets help humans solve multiclass classification tasks by narrowing down the set of potential label values to a subset of them, namely a prediction set, and asking them to always predict label values from the prediction sets. While this type of systems have been proven to be effective at improving the average accuracy of the predictions made by humans, by restricting human agency, they may cause harma human who has succeeded at predicting the ground-truth label of an instance on their own may have failed had they used these systems. In this paper, our goal is to control how frequently a decision support system based on prediction sets may cause harm, by design. To this end, we start by characterizing the above notion of harm using the theoretical framework of structural causal models. Then, we show that, under a natural, albeit unverifiable, monotonicity assumption, we can estimate how frequently a system may cause harm using only predictions made by humans on their own. Further, we also show that, under a weaker monotonicity assumption, which can be verified experimentally, we can bound how frequently a system may cause harm again using only predictions made by humans on their own. Building upon these assumptions, we introduce a computational framework to design decision support systems based on prediction sets that are guaranteed to cause harm less frequently than a user-specified value using conformal risk control. We validate our framework using real human predictions from two different human subject studies and show that, in decision support systems based on prediction sets, there is a trade-off between accuracy and counterfactual harm.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Human-AI Collaborative Uncertainty QuantificationSima Noorani, Shayan Kiyani, George Pappas, Hamed HassaniICML 2026 · 被引用 8 次
- Multi-Round Human–AI Collaboration with User-Specified RequirementsSima Noorani, Shayan Kiyani, Hamed Hassani, George PappasICML 2026 · 被引用 2 次
- Adaptively Grouped Contextual Bandits for Heterogeneous Human-AI Decision Making with Conformal Prediction SetsYanchen Wu, Bo LiICML 2026
- Beyond Accuracy: Latent Perturbations for Cognitive-Aware DiagnosisYuting Yan, Yinghao Fu, Wendi Ren, Haozhou Gao 等ICML 2026
- Conformal Prediction Sets Can Cause Disparate ImpactJesse C. Cresswell, Bhargava Kumar, Yi Sui, Mouloud BelbahriICLR 2025
它引用的顶会 Paper17
- Classification with Valid and Adaptive CoverageYaniv Romano, Matteo Sesia, Emmanuel J. CandèsNeurIPS 2020 · 被引用 586 次
- Consistent Estimators for Learning to Defer to an ExpertHussein Mozannar, David A. SontagICML 2020 · 被引用 267 次
- Conformal Risk ControlAnastasios Nikolas Angelopoulos, Stephen Bates, Adam Fisch, Lihua Lei 等ICLR 2024 · 被引用 242 次
- Differentiable Learning Under TriageNastaran Okati, Abir De, Manuel Gomez-RodriguezNeurIPS 2021 · 被引用 99 次
- Classification Under Human AssistanceAbir De, Nastaran Okati, Ali Zarezade, Manuel Gomez RodriguezAAAI 2021 · 被引用 60 次
相关 Paper
- Designing Decision Support Systems using Counterfactual Prediction SetsEleni Straitouri, Manuel Gomez RodriguezICML 2024 · 被引用 24 次
- Improving Expert Predictions with Conformal PredictionEleni Straitouri, Lequn Wang, Nastaran Okati, Manuel Gomez RodriguezICML 2023 · 被引用 56 次
- Towards Human-AI Complementarity with Prediction SetsGiovanni De Toni, Nastaran Okati, Suhas Thejaswi, Eleni Straitouri 等NeurIPS 2024
- Conformal Prediction Sets Improve Human Decision MakingJesse C. Cresswell, Yi Sui, Bhargava Kumar, Noël VouitsisICML 2024 · 被引用 36 次
- Optimal Decision-Making Based on Prediction SetsTao Wang, Edgar DobribanICML 2026
