Adaptively Grouped Contextual Bandits for Heterogeneous Human-AI Decision Making with Conformal Prediction Sets
Yanchen Wu, Bo Li
摘要
Personalizing AI decision support for heterogeneous human decision-makers remains a key challenge. We study a collaboration workflow where AI provides a reduced prediction set via conformal prediction and the human makes the final decision based on the set. We formulate this personalization problem as a contextual bandit, where individual and task features form the context, candidate significance levels serve as arms, and the optimal prediction-set size varies across contexts. To address large arm spaces and high-dimensional contexts, we introduce the Adaptively Grouped Contextual Bandit (AGCB) framework, which avoids global function approximation by exploiting two Human-AI structural assumptions: continuity and monotonicity. Continuity enables information sharing across nearby contexts and decisions, and drives a data-driven Zooming Mechanism that balances intra-group estimation error against inter-group approximation bias. Monotonicity converts each observation into directional counterfactual information over the candidate values, reducing the arm-dependence factor from polynomial to logarithmic in . Together, these mechanisms yield minimax-optimal dependence on the learning horizon for both cumulative and simple regret objectives. Empirical results confirm that AGCB achieves the strongest overall performance across most heterogeneous, data-scarce settings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper15
- Neural Contextual Bandits with UCB-based ExplorationDongruo Zhou, Lihong Li, Quanquan GuICML 2020 · 被引用 329 次
- Conformal Risk ControlAnastasios Nikolas Angelopoulos, Stephen Bates, Adam Fisch, Lihua Lei 等ICLR 2024 · 被引用 242 次
- Is the Most Accurate AI the Best Teammate? Optimizing AI for TeamworkGagan Bansal, Besmira Nushi, Ece Kamar, Eric Horvitz 等AAAI 2021 · 被引用 185 次
- Who Should I Trust: AI or Myself? Leveraging Human and AI Correctness Likelihood to Promote Appropriate Trust in AI-Assisted Decision-MakingShuai Ma, Ying Lei, Xinru Wang, Chengbo Zheng 等CHI 2023 · 被引用 139 次
- Model Selection in Contextual Stochastic Bandit ProblemsAldo Pacchiano, My Phan, Yasin Abbasi-Yadkori, Anup Rao 等NeurIPS 2020 · 被引用 107 次
相关 Paper
- Designing Decision Support Systems using Counterfactual Prediction SetsEleni Straitouri, Manuel Gomez RodriguezICML 2024 · 被引用 24 次
- Human-AI Collaborative Uncertainty QuantificationSima Noorani, Shayan Kiyani, George Pappas, Hamed HassaniICML 2026 · 被引用 8 次
- Proportional Response: Contextual Bandits for Simple and Cumulative Regret MinimizationSanath Kumar Krishnamurthy, Ruohan Zhan, Susan Athey, Emma BrunskillNeurIPS 2023 · 被引用 15 次
- Learning Personalized Decision Support PoliciesUmang Bhatt, Valerie Chen, Katherine M. Collins, Parameswaran Kamalaruban 等AAAI 2025 · 被引用 14 次
- Conformal Prediction Sets Improve Human Decision MakingJesse C. Cresswell, Yi Sui, Bhargava Kumar, Noël VouitsisICML 2024 · 被引用 36 次
