Observation Interference in Partially Observable Assistance Games
Scott Emmons, Caspar Oesterheld, Vincent Conitzer, Stuart Russell
2025年份
1顶会引用
摘要
We study partially observable assistance games (POAGs), a model of the human-AI value alignment problem which allows the human and the AI assistant to have partial observations. Motivated
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper6
- "Other-Play" for Zero-Shot CoordinationHengyuan Hu, Adam Lerer, Alex Peysakhovich, Jakob N. FoersterICML 2020 · 被引用 271 次
- Honesty Is the Best Policy: Defining and Mitigating AI DeceptionFrancis Ward, Francesca Toni, Francesco Belardinelli, Tom EverittNeurIPS 2023 · 被引用 60 次
- The Boltzmann Policy Distribution: Accounting for Systematic Suboptimality in Human ModelsCassidy Laidlaw, Anca D. DraganICLR 2022 · 被引用 46 次
- A New Formalism, Method and Open Issues for Zero-Shot CoordinationJohannes Treutlein, Michael Dennis, Caspar Oesterheld, Jakob N. FoersterICML 2021 · 被引用 45 次
- Learning to Interactively Learn and AssistMark Woodward, Chelsea Finn, Karol HausmanAAAI 2020 · 被引用 37 次
相关 Paper
- A decision-theoretic representation of assistive interfacesJulien Gori, Aurélien Nioche, Christoph Albert Johns, Antti OulasvirtaCHI 2026 · 被引用 1 次
- AssistanceZero: Scalably Solving Assistance GamesCassidy Laidlaw, Eli Bronstein, Timothy Guo, Dylan Feng 等ICML 2025
- Belief-Driven Value Alignment for Human-Robot CollaborationSaisai Li, Bing Shi, Yiming Xia, Xiao SuAAAI 2026
- Value Alignment VerificationDaniel S. Brown, Jordan Schneider, Anca D. Dragan, Scott NiekumICML 2021 · 被引用 41 次
- PAMDP: Interact to Persona Alignment via a Partially Observable Markov Decision ProcessZhe Yang, Yi Huang, Si Chen, Xiaoting Wu 等ICLR 2026
