"I Don't Think So": Summarizing Policy Disagreements for Agent Comparison
Yotam Amitai, Ofra Amir
Abstract
With Artificial Intelligence on the rise, human interaction with autonomous agents becomes more frequent. Effective human-agent collaboration requires users to understand the agent's behavior, as failing to do so may cause reduced productivity, misuse or frustration. Agent strategy summarization methods are used to describe the strategy of an agent to its destined user through demonstration. A summary's objective is to maximize the user's understanding of the agent's aptitude by showcasing its behaviour in a selected set of world states. While shown to be useful, we show that current methods are limited when tasked with comparing between agents, as each summary is independently generated for a specific agent. In this paper, we propose a novel method for generating dependent and contrastive summaries that emphasize the differences between agent policies by identifying states in which the agents disagree on the best course of action. We conduct user studies to assess the usefulness of disagreementbased summaries for identifying superior agents and conveying agent differences. Results show disagreement-based summaries lead to improved user performance compared to summaries generated using HIGHLIGHTS, a strategy summarization algorithm which generates summaries for each agent independently.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 32fcca49-a077-4518-83c1-77ad01dd87f8Cited by top-tier papers1
Ask how each one uses itRelated papers
- Explaining Decentralized Multi-Agent Reinforcement Learning PoliciesKayla Boggess, Sarit Kraus, Lu FengAAAI 2026
- Contrastive Explanations That Anticipate Human Misconceptions Can Improve Human Decision-Making SkillsZana Buçinca, Siddharth Swaroop, Amanda E. Paluch, Finale Doshi-Velez et al.CHI 2025 · 31 citations
- An Evaluation of Situational Autonomy for Human-AI Collaboration in a Shared Workspace SettingVildan Salikutluk, Janik Schöpper, Franziska Herbert, Katrin Scheuermann et al.CHI 2024 · 27 citations
- Advancing Collaborative Debates with Role Differentiation through Multi-Agent Reinforcement LearningHaoran Li, Ziyi Su, Yun Xue, Zhiliang Tian et al.ACL 2025 · 8 citations
- Other Roles Matter! Enhancing Role-Oriented Dialogue Summarization via Role InteractionsHaitao Lin, Junnan Zhu, Lu Xiang, Yu Zhou et al.ACL 2022 · 36 citations
