From Stimuli to Minds: Enhancing Psychological Reasoning in LLMs via Bilateral Reinforcement Learning
Yichao Feng, Haoran Luo, Lang Feng, Shuai Zhao, Anh Tuan Luu
摘要
Large Language Models show promise in emotion understanding, social reasoning, and empathy, yet they struggle with psychologically grounded tasks that require inferring implicit mental states in context-rich, ambiguous settings. These limitations arise from the absence of theory-aligned supervision and the difficulty of capturing nuanced mental processes in real-world narratives. To address this gap, we leverage expert-labeled, psychologically rich scenarios and propose a trajectory-aware reinforcement learning framework that explicitly imitates expert psychological thought patterns. By integrating real-world stimuli with structured reasoning guidance, our approach enables compact models to internalize social-cognitive principles, perform nuanced psychological inference, and support continual self-improvement. Experiments across multiple benchmarks demonstrate that our models achieve expert-level interpretive capabilities, exhibiting strong out-of-distribution generalization across diverse psychologically tasks. Our code is publicly available at https://github.com/Githubuseryf/Stimuli2Minds . * These authors contributed equally. Mrs. Li's son, Xiao Qiang, is in college now, he tells her that he is the captain of the school football team. This achievement is high because the football team is very popular in school. Question: How does Mrs. Li react when Xiao Qiang is chosen as the football team captain? Mrs. Li's son, Xiaoqiang, goes to college and tells her that he is the captain of the school football team. This achievement is high because the football team is very popular in school. But Mrs. Li thinks her son may face pressure in balancing study and football activities. Xiaoqiang tells Mrs. Li that he is the captain of the football team. Question: What emotion does Mrs. Li possibly show? Mrs. Li would likely be very proud of her son for being chosen as the football team captain
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper8
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- Group-in-Group Policy Optimization for LLM Agent TrainingLang Feng, Zhenghai Xue, Tingcong Liu, Bo AnNeurIPS 2025 · 被引用 484 次
- CogBench: a large language model walks into a psychology labJulian Coda-Forno, Marcel Binz, Jane X. Wang, Eric SchulzICML 2024 · 被引用 60 次
- PsyDT: Using LLMs to Construct the Digital Twin of Psychological Counselor with Personalized Counseling Style for Psychological CounselingHaojie Xie, Yirong Chen, Xiaofen Xing, Jingkai Lin 等ACL 2025 · 被引用 43 次
- SimpleToM: Exposing the Gap between Explicit ToM Inference and Implicit ToM Application in LLMsYuling Gu, Oyvind Tafjord, Hyunwoo Kim, Jared Moore 等ICLR 2026 · 被引用 39 次
相关 Paper
- EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language ModelsYiyang Fang, Wenke Huang, Pei Fu, Yihao Yang 等CVPR 2026 · 被引用 4 次
- Unveiling the Cognitive Compass: Theory-of-Mind-Guided Multimodal Emotion ReasoningMeng Luo, Bobo Li, Shanqing Xu, Shize Zhang 等ICLR 2026 · 被引用 10 次
- MetaMind: Modeling Human Social Thoughts with Metacognitive Multi-Agent SystemsXuanming Zhang, Yuxuan Chen, Samuel (Min-Hsuan) Yeh, Sharon LiNeurIPS 2025 · 被引用 14 次
- EmotionThinker: Prosody-Aware Reinforcement Learning for Explainable Speech Emotion ReasoningDingdong WANG, Shujie LIU, Tianhua Zhang, Youjun Chen 等ICLR 2026 · 被引用 22 次
- Modeling Protagonist Emotions for Emotion-Aware StorytellingFaeze Brahman, Snigdha ChaturvediEMNLP 2020 · 被引用 5 次
