Know You First and Be You Better: Modeling Human-Like User Simulators via Implicit Profiles
Kuang Wang, Xianfei Li, Shenghao Yang, Li Zhou, Feng Jiang, Haizhou Li
摘要
User simulators are crucial for replicating human interactions with dialogue systems, supporting both collaborative training and automatic evaluation, especially for large language models (LLMs). However, current role-playing methods face challenges such as a lack of utterance-level authenticity and user-level diversity, often hindered by role confusion and dependence on predefined profiles of wellknown figures. In contrast, direct simulation focuses solely on text, neglecting implicit user traits like personality and conversation-level consistency. To address these issues, we introduce the User Simulator with Implicit Profiles (USP), a framework that infers implicit user profiles from human-machine interactions to simulate personalized and realistic dialogues. We first develop an LLM-driven extractor with a comprehensive profile schema, then refine the simulation using conditional supervised finetuning and reinforcement learning with cycle consistency, optimizing at both the utterance and conversation levels. Finally, a diverse profile sampler captures the distribution of realworld user profiles. Experimental results show that USP outperforms strong baselines in terms of authenticity and diversity while maintaining comparable consistency. Additionally, using USP to evaluate LLM on dynamic multiturn aligns well with mainstream benchmarks, demonstrating its effectiveness in real-world applications. We open-source related resources in https://github.com/wangkevin02/USP .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Flipping the Dialogue: Training and Evaluating User Language ModelsTarek Naous, Philippe Laban, Wei Xu, Jennifer NevilleICLR 2026 · 被引用 56 次
- HAG: Hierarchical Demographic Tree-based Agent Generation for Topic-Adaptive SimulationRongxin Chen, Tianyu Wu, Bingbing Xu, Jiatang Luo 等ACL 2026 · 被引用 1 次
- HumanLM: Simulating Users with State Alignment Beats Response ImitationShirley Wu, Evelyn Choi, Arpandeep Khatua, Zhanghan Wang 等ICML 2026
- Unveiling and Simulating Short-Video Addiction Behaviors via Economic Addiction TheoryChen Xu, Zhipeng Yi, Ruizi Wang, Wenjie Wang 等WWW 2026
- DiscoverLLM: From Executing Intents to Discovering ThemTae Soo Kim, Yoonjoo Lee, Jaesang Yu, John Chung 等ICML 2026
它引用的顶会 Paper22
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 被引用 2,496 次
- Understanding Contrastive Representation Learning through Alignment and Uniformity on the HypersphereTongzhou Wang, Phillip IsolaICML 2020 · 被引用 2,360 次
- Chatbot Arena: An Open Platform for Evaluating LLMs by Human PreferenceWei-Lin Chiang, Lianmin Zheng, Ying Sheng, Anastasios Nikolas Angelopoulos 等ICML 2024 · 被引用 1,212 次
- LMSYS-Chat-1M: A Large-Scale Real-World LLM Conversation DatasetLianmin Zheng, Wei-Lin Chiang, Ying Sheng, Tianle Li 等ICLR 2024 · 被引用 419 次
- FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text GenerationSewon Min, Kalpesh Krishna, Xinxi Lyu, Mike Lewis 等EMNLP 2023 · 被引用 225 次
相关 Paper
- Consistently Simulating Human Personas with Multi-Turn Reinforcement LearningMarwa Abdulhai, Ryan Cheng, Donovan Clay, Tim Althoff 等NeurIPS 2025 · 被引用 51 次
- Diagnostic-Guided Dynamic Profile Optimization for LLM-based User Simulators in Sequential RecommendationHongyang Liu, Zhu Sun, Tianjun Wei, Yan Wang 等AAAI 2026 · 被引用 4 次
- AlignUSER: Human-Aligned LLM Agents via World Models for Recommender System EvaluationNicolas Bougie, Gian Maria Marconi, Xiaotong Ye, Narimawa WatanabeACL 2026 · 被引用 3 次
- Teaching Language Models to Evolve with Users: Dynamic Profile Modeling for Personalized AlignmentWeixiang Zhao, Xingyu Sui, Yulin Hu, Jiahe Guo 等NeurIPS 2025 · 被引用 34 次
- Task-Aware Automated User Profile Generation for Recommendation Simulation Using Large Language ModelsXinye Wanyan, Chenglong Ma, Danula Hettiachchi, Ziqi Xu 等SIGIR 2026
