Shared Autonomy with IDA: Interventional Diffusion Assistance
Brandon McMahan, Zhenghao Mark Peng, Bolei Zhou, Jonathan C. Kao
Abstract
The rapid development of artificial intelligence (AI) has unearthed the potential to assist humans in controlling advanced technologies. Shared autonomy (SA) facilitates control by combining inputs from a human pilot and an AI copilot. In prior SA studies, the copilot is constantly active in determining the action played at each time step. This limits human autonomy and may have deleterious effects on performance. In general, the amount of helpful copilot assistance can vary greatly depending on the task dynamics. We therefore hypothesize that human autonomy and SA performance improve through dynamic and selective copilot intervention. To address this, we develop a goal-agnostic intervention assistance (IA) that dynamically shares control by having the copilot intervene only when the expected value of the copilot's action exceeds that of the human's action across all possible goals. We implement IA with a diffusion copilot (termed IDA) trained on expert demonstrations with goal masking. We prove a lower bound on the performance of IA that depends on pilot and copilot performance. Experiments with simulated human pilots show that IDA achieves higher performance than pilot-only and traditional SA control in variants of the Reacher environment and Lunar Lander. We then demonstrate that IDA achieves better control in Lunar Lander with human-in-the-loop experiments. Human participants report greater autonomy with IDA and prefer IDA over pilot-only and traditional SA control. We attribute the success of IDA to preserving human autonomy while simultaneously offering assistance to prevent the human pilot from entering universally bad states.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e3db5d7b-1d54-4848-81e4-2b1add58b58dBuilds on6
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Efficient Learning of Safe Driving Policy via Human-AI Copilot OptimizationQuanyi Li, Zhenghao Peng, Bolei ZhouICLR 2022 · 80 citations
- Learning from Active Human Involvement through Proxy Value PropagationZhenghao Mark Peng, Wenjie Mo, Chenda Duan, Quanyi Li et al.NeurIPS 2023 · 30 citations
- Randomized Ensembled Double Q-Learning: Learning Fast Without a ModelXinyue Chen, Che Wang, Zijian Zhou, Keith W. RossICLR 2021 · 26 citations
- On Optimizing Interventions in Shared AutonomyWeihao Tan, David Koleczek, Siddhant Pradhan, Nicholas Perello et al.AAAI 2022 · 6 citations
Related papers
- Robot-Gated Interactive Imitation Learning with Adaptive Intervention MechanismHaoyuan Cai, Zhenghao Peng, Bolei ZhouICML 2025
- Two Heads Are Better Than One: A Dimension Space for Unifying Human and Artificial Intelligence in Shared ControlGabriele Cimolino, T. C. Nicholas GrahamCHI 2022 · 15 citations
- An Evaluation of Situational Autonomy for Human-AI Collaboration in a Shared Workspace SettingVildan Salikutluk, Janik Schöpper, Franziska Herbert, Katrin Scheuermann et al.CHI 2024 · 27 citations
- AvE: Assistance via EmpowermentYuqing Du, Stas Tiomkin, Emre Kiciman, Daniel Polani et al.NeurIPS 2020 · 51 citations
- Policy Optimization under Imperfect Human Interactions with Agent-Gated Shared AutonomyZhenghai Xue, Bo An, Shuicheng YanICLR 2025
