Reducing Goal State Divergence with Environment Design
Kelsey Sikes, Sarah Keren, Sarath Sreedharan
摘要
Generating behaviors that align with human expectations is a key requirement for human-robot collaboration. Potential behavior misalignment could lead to the robot performing actions with unanticipated, potentially dangerous side effects even while pursuing human goals. In this paper, we introduce a novel metric called Goal State Divergence (GSD) which quantifies the difference between the state a robot achieved in response to a human-specified goal and what the human expected. In cases where GSD cannot be directly calculated, we show how it can be approximated using maximal and minimal bounds. We then leverage GSD in our novel human-robot goal alignment design (HRGAD) problem, which identifies a minimal set of environment modifications that can reduce such mismatches. We show the effectiveness of our method in reducing the goal state divergence by empirically evaluating our approach on several planning benchmarks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper2
相关 Paper
- f-Policy Gradients: A General Framework for Goal-Conditioned RL using f-DivergencesSiddhant Agarwal, Ishan Durugkar, Peter Stone, Amy ZhangNeurIPS 2023 · 被引用 22 次
- Expectation Alignment: Handling Reward Misspecification in the Presence of Expectation MismatchMalek Mechergui, Sarath SreedharanNeurIPS 2024 · 被引用 4 次
- Inferring Implicit Goals Across Differing Task ModelsSilvia Tulli, Stylianos Loukas Vasileiou, Mohamed Chetouani, Sarath SreedharanAAAI 2026
- Extended Goal Recognition Design with First-Order Computation Tree LogicTsz-Chiu AuAAAI 2022 · 被引用 2 次
- Stochastic Goal Recognition Design Problems with Suboptimal AgentsChristabel Wayllace, William YeohAAAI 2022 · 被引用 3 次
