Evaluating Character Understanding of Large Language Models via Character Profiling from Fictional Works
Xinfeng Yuan, Siyu Yuan, Yuhan Cui, Tianhe Lin, Xintao Wang, Rui Xu, Jiangjie Chen, Deqing Yang
摘要
Large language models (LLMs) have demonstrated impressive performance and spurred numerous AI applications, in which role-playing agents (RPAs) are particularly popular, especially for fictional characters. The prerequisite for these RPAs lies in the capability of LLMs to understand characters from fictional works. Previous efforts have evaluated this capability via basic classification tasks or characteristic imitation, failing to capture the nuanced character understanding with LLMs. In this paper, we propose evaluating LLMs' character understanding capability via the character profiling task, i.e., summarizing character profiles from corresponding materials, a widely adopted yet understudied practice for RPA development. Specifically, we construct the CROSS dataset from literature experts and assess the generated profiles by comparing them with ground truth references and evaluating their applicability in downstream tasks. Our experiments, which cover various summarization methods and LLMs, have yielded promising results. These results strongly validate the character understanding capability of LLMs. Resources are available at https://github. com/Joanna0123/character_profiling .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Revealing and Mitigating the Challenge of Detecting Character Knowledge Errors in LLM Role-PlayingWenyuan Zhang, Shuaiyi Nie, Jiawei Sheng, Zefeng Zhang 等EMNLP 2025 · 被引用 10 次
- BOOKWORLD: From Novels to Interactive Agent Societies for Story CreationYiting Ran, Xintao Wang, Tian Qiu, Jiaqing Liang 等ACL 2025 · 被引用 9 次
- Codifying Character Logic in Role-PlayingLetian Peng, Jingbo ShangNeurIPS 2025 · 被引用 7 次
- RolePlot: A Systematic Framework for Evaluating and Enhancing the Plot-Progression Capabilities of Role-Playing AgentsPinyi Zhang, Siyu An, Lingfeng Qiao, Yifei Yu 等ACL 2025 · 被引用 4 次
- DEEPER Insight into Your User: Directed Persona Refinement for Dynamic Persona ModelingAili Chen, Chengyu Du, Jiangjie Chen, Jinghan Xu 等ACL 2025
它引用的顶会 Paper9
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger 等ICLR 2020 · 被引用 8,443 次
- BooookScore: A systematic exploration of book-length summarization in the era of LLMsYapei Chang, Kyle Lo, Tanya Goyal, Mohit IyyerICLR 2024 · 被引用 173 次
- Character-LLM: A Trainable Agent for Role-PlayingYunfan Shao, Linyang Li, Junqi Dai, Xipeng QiuEMNLP 2023 · 被引用 97 次
- Translate Meanings, Not Just Words: IdiomKB's Role in Optimizing Idiomatic Translation with Language ModelsShuang Li, Jiangjie Chen, Siyu Yuan, Xinyi Wu 等AAAI 2024 · 被引用 44 次
- Few-Shot Character Understanding in Movies as an Assessment to Meta-Learning of Theory-of-MindMo Yu, Qiujing Wang, Shunchi Zhang, Yisi Sang 等ICML 2024 · 被引用 22 次
相关 Paper
- Video2Roleplay: A Multimodal Dataset and Framework for Video-Guided Role-playing AgentsXueqiao Zhang, Chao Zhang, Jingtao Xu, Yifan Zhu 等EMNLP 2025 · 被引用 2 次
- InCharacter: Evaluating Personality Fidelity in Role-Playing Agents through Psychological InterviewsXintao Wang, Yunze Xiao, Jen-tse Huang, Siyu Yuan 等ACL 2024
- Speaker Verification in Agent-generated ConversationsYizhe Yang, Palakorn Achananuparp, Heyan Huang, Jing Jiang 等ACL 2024
- CoSER: Coordinating LLM-Based Persona Simulation of Established RolesXintao Wang, Heng Wang, Yifei Zhang, Xinfeng Yuan 等ICML 2025
- Thinking in Character: Advancing Role-Playing Agents with Role-Aware ReasoningYihong Tang, Kehai Chen, Muyun Yang, Zheng-Yu Niu 等NeurIPS 2025 · 被引用 16 次
