Confounding-Robust Deferral Policy Learning
Ruijiang Gao, Mingzhang Yin
Abstract
Human-AI collaboration has the potential to transform various domains by leveraging the complementary strengths of human experts and Artificial Intelligence (AI) systems. However, unobserved confounding can undermine the effectiveness of this collaboration, leading to biased and unreliable outcomes. In this paper, we propose a novel solution to address unobserved confounding in human-AI collaboration by employing sensitivity analysis from causal inference. Our approach combines domain expertise with AI-driven statistical modeling to account for potentially hidden confounders. We present a deferral collaboration framework for incorporating the sensitivity model into offline policy learning, enabling the system to control for the influence of unobserved confounding factors. In addition, we propose a personalized deferral collaboration system to leverage the diverse expertise of different human decision-makers. By adjusting for potential biases, our proposed solution enhances the robustness and reliability of collaborative outcomes. The empirical and theoretical analyses demonstrate the efficacy of our approach in mitigating unobserved confounding and improving the overall performance of human-AI collaborations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ba47e789-ce3f-463a-b097-bd0a027994c7Cited by top-tier papers3
- Deferring Concept Bottleneck Models: Learning to Defer Interventions to Inaccurate ExpertsAndrea Pugnana, Riccardo Massidda, Francesco Giannini, Pietro Barbiero et al.NeurIPS 2025 · 11 citations
- A Causal Target for Learning to Defer Under Hidden ConfoundingYanmin Li, Lihua Liu, Xin Wang, Zhilong Mao et al.AAAI 2026
- Treatment Responder Classification with AbstentionHaoxiang Wang, Haoxuan Li, Ziyan Wang, Zhiheng Zhang et al.ICML 2026
Builds on2
- Regression under Human AssistanceAbir De, Paramita Koley, Niloy Ganguly, Manuel Gomez-RodriguezAAAI 2020 · 73 citations
- Toward Supporting Perceptual Complementarity in Human-AI Collaboration via Reflection on UnobservablesKenneth Holstein, Maria De-Arteaga, Lakshmi Tumati, Yanghuidi ChengCSCW 2023 · 37 citations
Related papers
- When to Act and When to Ask: Policy Learning With Deferral Under Hidden ConfoundingMarah Ghoummaid, Uri ShalitNeurIPS 2024 · 4 citations
- Efficient and Sharp Off-Policy Learning under Unobserved ConfoundingKonstantin Hess, Dennis Frauen, Valentyn Melnychuk, Stefan FeuerriegelICLR 2026 · 5 citations
- Probabilistic Learning to Defer: Handling Missing Expert Annotations and Controlling Workload DistributionCuong C. Nguyen, Thanh-Toan Do, Gustavo CarneiroICLR 2025
- Confounding Robust Deep Reinforcement Learning: A Causal ApproachMingxuan Li, Junzhe Zhang, Elias BareinboimNeurIPS 2025 · 7 citations
- Understanding Choice Independence and Error Types in Human-AI CollaborationAlexander Erlei, Abhinav Sharma, Ujwal GadirajuCHI 2024 · 25 citations
