A Causal Target for Learning to Defer Under Hidden Confounding
Yanmin Li, Lihua Liu, Xin Wang, Zhilong Mao, Jibing Wu, Weidong Bao
Abstract
Learning decision policies from confounded observational data is a challenging task in causal inference, as unobserved confounders can lead to biased or suboptimal actions when relying solely on machine learning models. A synergistic approach is learning to defer, which decides when to act itself and when to defer to a human expert with access to unobserved information. However, constructing the learning target, which defines the probability of choosing each action or deferral, remains a core challenge. To address this, we propose causal-target-based learning to defer (CTLD) framework, where the causal target is constructed from sharp bounds on potential outcomes. Specifically, the degree of overlap between these bounds determines the probability of deferral, while their relative positions and widths define the probabilities over actions. CTLD aligns model predictions with this causal target to make probabilistic decisions over actions and deferral. We present comprehensive theoretical guarantees for the learned policy and demonstrate the effectiveness of CTLD on synthetic and semi-synthetic datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 29237d37-33f9-4254-9202-3c0e5eac912eBuilds on8
- Quantifying Ignorance in Individual-Level Causal-Effect Estimates under Hidden ConfoundingAndrew Jesson, Sören Mindermann, Yarin Gal, Uri ShalitICML 2021 · 66 citations
- Role of Human-AI Interaction in Selective PredictionElizabeth Bondi, Raphael Koster, Hannah Sheahan, Martin J. Chadwick et al.AAAI 2022 · 46 citations
- B-Learner: Quasi-Oracle Bounds on Heterogeneous Causal Effects Under Hidden ConfoundingMiruna Oprescu, Jacob Dorn, Marah Ghoummaid, Andrew Jesson et al.ICML 2023 · 39 citations
- In Defense of Softmax Parametrization for Calibrated and Consistent Learning to DeferYuzhou Cao, Hussein Mozannar, Lei Feng, Hongxin Wei et al.NeurIPS 2023 · 36 citations
- Learning to Defer with Limited Expert PredictionsPatrick Hemmer, Lukas Thede, Michael Vössing, Johannes Jakubik et al.AAAI 2023 · 28 citations
Related papers
- When to Act and When to Ask: Policy Learning With Deferral Under Hidden ConfoundingMarah Ghoummaid, Uri ShalitNeurIPS 2024 · 4 citations
- Deferring Concept Bottleneck Models: Learning to Defer Interventions to Inaccurate ExpertsAndrea Pugnana, Riccardo Massidda, Francesco Giannini, Pietro Barbiero et al.NeurIPS 2025 · 11 citations
- Confounding-Robust Deferral Policy LearningRuijiang Gao, Mingzhang YinAAAI 2025 · 2 citations
- Exploiting Human-AI Dependence for Learning to DeferZixi Wei, Yuzhou Cao, Lei FengICML 2024 · 15 citations
- Towards Safe Policy Learning under Partial Identifiability: A Causal ApproachShalmali Joshi, Junzhe Zhang, Elias BareinboimAAAI 2024 · 10 citations
