Confounding-Robust Deferral Policy Learning
Ruijiang Gao, Mingzhang Yin
摘要
Human-AI collaboration has the potential to transform various domains by leveraging the complementary strengths of human experts and Artificial Intelligence (AI) systems. However, unobserved confounding can undermine the effectiveness of this collaboration, leading to biased and unreliable outcomes. In this paper, we propose a novel solution to address unobserved confounding in human-AI collaboration by employing sensitivity analysis from causal inference. Our approach combines domain expertise with AI-driven statistical modeling to account for potentially hidden confounders. We present a deferral collaboration framework for incorporating the sensitivity model into offline policy learning, enabling the system to control for the influence of unobserved confounding factors. In addition, we propose a personalized deferral collaboration system to leverage the diverse expertise of different human decision-makers. By adjusting for potential biases, our proposed solution enhances the robustness and reliability of collaborative outcomes. The empirical and theoretical analyses demonstrate the efficacy of our approach in mitigating unobserved confounding and improving the overall performance of human-AI collaborations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Deferring Concept Bottleneck Models: Learning to Defer Interventions to Inaccurate ExpertsAndrea Pugnana, Riccardo Massidda, Francesco Giannini, Pietro Barbiero 等NeurIPS 2025 · 被引用 11 次
- A Causal Target for Learning to Defer Under Hidden ConfoundingYanmin Li, Lihua Liu, Xin Wang, Zhilong Mao 等AAAI 2026
- Treatment Responder Classification with AbstentionHaoxiang Wang, Haoxuan Li, Ziyan Wang, Zhiheng Zhang 等ICML 2026
它引用的顶会 Paper2
- Regression under Human AssistanceAbir De, Paramita Koley, Niloy Ganguly, Manuel Gomez-RodriguezAAAI 2020 · 被引用 73 次
- Toward Supporting Perceptual Complementarity in Human-AI Collaboration via Reflection on UnobservablesKenneth Holstein, Maria De-Arteaga, Lakshmi Tumati, Yanghuidi ChengCSCW 2023 · 被引用 37 次
相关 Paper
- When to Act and When to Ask: Policy Learning With Deferral Under Hidden ConfoundingMarah Ghoummaid, Uri ShalitNeurIPS 2024 · 被引用 4 次
- Efficient and Sharp Off-Policy Learning under Unobserved ConfoundingKonstantin Hess, Dennis Frauen, Valentyn Melnychuk, Stefan FeuerriegelICLR 2026 · 被引用 5 次
- Probabilistic Learning to Defer: Handling Missing Expert Annotations and Controlling Workload DistributionCuong C. Nguyen, Thanh-Toan Do, Gustavo CarneiroICLR 2025
- Confounding Robust Deep Reinforcement Learning: A Causal ApproachMingxuan Li, Junzhe Zhang, Elias BareinboimNeurIPS 2025 · 被引用 7 次
- Understanding Choice Independence and Error Types in Human-AI CollaborationAlexander Erlei, Abhinav Sharma, Ujwal GadirajuCHI 2024 · 被引用 25 次
