METRO: Towards Strategy Induction from Expert Dialogue Transcripts for Non-collaborative Dialogues
Haofu Yang, Jiaji Liu, Chen Huang, Faguo Wu, Wenqiang Lei, See-Kiong Ng
摘要
Developing non-collaborative dialogue agents traditionally requires the manual, unscalable codification of expert strategies. We propose METRO, a method that leverages large language models to autonomously induce both strategy actions and planning logic directly from raw transcripts. METRO formalizes expert knowledge into a Strategy Forest, a hierarchical structure that captures both shortterm responses (nodes) and long-term strategic foresight (branches). Experimental results across two benchmarks show that METRO demonstrates promising performance, outperforming existing methods by an average of 9%-10%. Our further analysis not only reveals the success behind METRO (strategic behavioral diversity and foresight), but also demonstrates its robust cross-task transferability. This offers new insights into building non-collaborative agents in a cost-effective and scalable way. Our code is available at https: //github.com/Humphrey-0125/METRO .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper14
- Evaluating and Inducing Personality in Pre-trained Language ModelsGuangyuan Jiang, Manjie Xu, Song-Chun Zhu, Wenjuan Han 等NeurIPS 2023 · 被引用 192 次
- Plug-and-Play Policy Planner for Large Language Model Powered Dialogue AgentsYang Deng, Wenxuan Zhang, Wai Lam, See-Kiong Ng 等ICLR 2024 · 被引用 86 次
- Roleplay-doh: Enabling Domain-Experts to Create LLM-simulated Patients via Eliciting and Adhering to PrinciplesRyan Louie, Ananjan Nandi, William Fang, Cheng Chang 等EMNLP 2024 · 被引用 37 次
- Augmenting Non-Collaborative Dialog Systems with Explicit Semantic and Strategic Dialog HistoryYiheng Zhou, Yulia Tsvetkov, Alan W. Black, Zhou YuICLR 2020 · 被引用 35 次
- Simulation-Free Hierarchical Latent Policy Planning for Proactive DialoguesTao He, Lizi Liao, Yixin Cao, Yuanxing Liu 等AAAI 2025 · 被引用 11 次
相关 Paper
- EPO: Explicit Policy Optimization for Strategic Reasoning in LLMs via Reinforcement LearningXiaoqian Liu, Ke Wang, Yongbin Li, Yuchuan Wu 等ACL 2025 · 被引用 7 次
- Inductive-Deductive Strategy Reuse for Multi-Turn Instructional DialoguesJiao Ou, Jiayu Wu, Che Liu, Fuzheng Zhang 等EMNLP 2024 · 被引用 2 次
- ChatSOP: An SOP-Guided MCTS Planning Framework for Controllable LLM Dialogue AgentsZhigen Li, Jianxiang Peng, Yanmeng Wang, Yong Cao 等ACL 2025 · 被引用 9 次
- DialogXpert: Driving Intelligent and Emotion-Aware Conversations Through Online Value-Based Reinforcement Learning with LLM PriorsTazeek Bin Abdur Rakib, Ambuj Mehrish, Lay-Ki Soon, Wern Han Lim 等AAAI 2026 · 被引用 3 次
- META: Meta Evolution of Tool Trajectory Adaptation for Long-Video UnderstandingJing Huang, Luyuan Chen, Zhijie Xu, Yadong Li 等CVPR 2026
