Persuading across Diverse Domains: a Dataset and Persuasion Large Language Model
Chuhao Jin, Kening Ren, Lingzhen Kong, Xiting Wang, Ruihua Song, Huan Chen
摘要
Persuasive dialogue requires multi-turn following and planning abilities to achieve the goal of persuading users, which is still challenging even for state-of-the-art large language models (LLMs). Previous works focus on retrievalbased models or generative models in a specific domain due to a lack of data across multiple domains. In this paper, we leverage GPT-4 to create the first multi-domain persuasive dialogue dataset DailyPersuasion. Then we propose a general method named PersuGPT to learn a persuasion model based on LLMs through intent-to-strategy reasoning, which summarizes the intent of user's utterance and reasons next strategy to respond. Moreover, we design a simulation-based preference optimization, which utilizes a learned user model and our model to simulate next turns and estimate their rewards more accurately. Experimental results on two datasets indicate that our proposed method outperforms all baselines in terms of automatic evaluation metric Win-Rate and human evaluation. The code and data are available at https://persugpt.github.io . Write a conversation with annotated intent-to-strategy reasoning processes, from a storytelling view, with the rules as follows… Scenario Example Keywords: Fundraising Background: During a volunteer activity, Alex is raising funds for animal protection and hopes her friend Mia can donate. Goal: Persuade Mia to donate. Keywords for New Scenarios 1. Cooking & Stress relief 2. Pet care 3. Social justice 4. digital economy ••• Refer to the scenario example, write new scenarios by the keywords, with the rules as follows… Guideline 1. Emotional factors 2. Social proof principle 3. Scarcity principle ••• Refer to the guideline, write strategies for the task, with the rules as follows… Prompt Keywords: Cooking & Stress relief Background: Mia notices her roommate Alex stressed from work, opts for instant noodles nightly to save money. Goal: Persuade Alex cooking can be a budget-friendly stress reliever Generated Scenarios 𝑪 𝒊 GPT-4 Intent to Strategy Reasoning: Alex's goal to save money and reduce stress guides the choice of Budget-friendly recipes strategy. It's crucial to connect on a personal level, showing understanding and offering a practical solution. Response: Hey Alex, let's cook together. It's cheaper than instant meals and a fun way to unwind. What do you say? Cooking seems pricey and time-consuming. Will it really save money?
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Evaluating Text Creativity across Diverse Domains: a Dataset and Large Language Model EvaluatorQian Cao, Xiting Wang, Yuzhuo Yuan, Yahui Liu 等ICLR 2026 · 被引用 11 次
- ToMAP: Training Opponent-Aware LLM Persuaders with Theory of MindPeixuan Han, Zijia Liu, Jiaxuan YouICML 2026 · 被引用 9 次
- OpenDeception: Learning Deception and Trust in Human–AI Interaction via Multi-Agent SimulationYichen Wu, Qianqian Gao, Xudong Pan, Geng Hong 等ICML 2026 · 被引用 1 次
- SQLWOZ: A Realistic Task-Oriented Dialogue Dataset with SQL-Based Dialogue State Representation for Complex User RequirementsHeng-Da Xu, Xian-Ling Mao, Fanshu Sun, Tian-Yi Che 等EMNLP 2025
- Data Speaks, But who Gives It a Voice? Understanding Persuasive Strategies in Data-Driven News ArticlesZikai Li, Chuyi Zheng, Ziang Li, Yang ShiIEEE VIS 2025
它引用的顶会 Paper14
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning 等NeurIPS 2023 · 被引用 10,924 次
- Multitask Prompted Training Enables Zero-Shot Task GeneralizationVictor Sanh, Albert Webson, Colin Raffel, Stephen H. Bach 等ICLR 2022 · 被引用 1,976 次
相关 Paper
- Evaluating Intention Detection Capability of Large Language Models in Persuasive DialoguesHiromasa Sakurai, Yusuke MiyaoACL 2024
- Measuring And Improving Persuasiveness Of Large Language ModelsSomesh Kumar Singh, Yaman Kumar Singla, Harini S. I, Balaji KrishnamurthyICLR 2025
- Weakly-Supervised Hierarchical Models for Predicting Persuasive Strategies in Good-faith Textual RequestsJiaao Chen, Diyi YangAAAI 2021 · 被引用 25 次
- Persuasion Dynamics in LLMs: Investigating Robustness and Adaptability in Knowledge and Safety with DuET-PDBryan Chen Zhengyu Tan, Daniel Wai Kit Chin, Zhengyuan Liu, Nancy F. Chen 等EMNLP 2025
- AI-Salesman: Towards Reliable Large Language Model Driven TelemarketingQingyu Zhang, Chunlei Xin, Xuanang Chen, Yaojie Lu 等AAAI 2026
