From Stimuli to Minds: Enhancing Psychological Reasoning in LLMs via Bilateral Reinforcement Learning
Yichao Feng, Haoran Luo, Lang Feng, Shuai Zhao, Anh Tuan Luu
Abstract
Large Language Models show promise in emotion understanding, social reasoning, and empathy, yet they struggle with psychologically grounded tasks that require inferring implicit mental states in context-rich, ambiguous settings. These limitations arise from the absence of theory-aligned supervision and the difficulty of capturing nuanced mental processes in real-world narratives. To address this gap, we leverage expert-labeled, psychologically rich scenarios and propose a trajectory-aware reinforcement learning framework that explicitly imitates expert psychological thought patterns. By integrating real-world stimuli with structured reasoning guidance, our approach enables compact models to internalize social-cognitive principles, perform nuanced psychological inference, and support continual self-improvement. Experiments across multiple benchmarks demonstrate that our models achieve expert-level interpretive capabilities, exhibiting strong out-of-distribution generalization across diverse psychologically tasks. Our code is publicly available at https://github.com/Githubuseryf/Stimuli2Minds . * These authors contributed equally. Mrs. Li's son, Xiao Qiang, is in college now, he tells her that he is the captain of the school football team. This achievement is high because the football team is very popular in school. Question: How does Mrs. Li react when Xiao Qiang is chosen as the football team captain? Mrs. Li's son, Xiaoqiang, goes to college and tells her that he is the captain of the school football team. This achievement is high because the football team is very popular in school. But Mrs. Li thinks her son may face pressure in balancing study and football activities. Xiaoqiang tells Mrs. Li that he is the captain of the football team. Question: What emotion does Mrs. Li possibly show? Mrs. Li would likely be very proud of her son for being chosen as the football team captain
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f9662e9b-1eac-47f0-b38c-0f0d2986936fCited by top-tier papers1
Ask how each one uses itBuilds on8
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Group-in-Group Policy Optimization for LLM Agent TrainingLang Feng, Zhenghai Xue, Tingcong Liu, Bo AnNeurIPS 2025 · 484 citations
- CogBench: a large language model walks into a psychology labJulian Coda-Forno, Marcel Binz, Jane X. Wang, Eric SchulzICML 2024 · 60 citations
- PsyDT: Using LLMs to Construct the Digital Twin of Psychological Counselor with Personalized Counseling Style for Psychological CounselingHaojie Xie, Yirong Chen, Xiaofen Xing, Jingkai Lin et al.ACL 2025 · 43 citations
- SimpleToM: Exposing the Gap between Explicit ToM Inference and Implicit ToM Application in LLMsYuling Gu, Oyvind Tafjord, Hyunwoo Kim, Jared Moore et al.ICLR 2026 · 39 citations
Related papers
- EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language ModelsYiyang Fang, Wenke Huang, Pei Fu, Yihao Yang et al.CVPR 2026 · 4 citations
- Unveiling the Cognitive Compass: Theory-of-Mind-Guided Multimodal Emotion ReasoningMeng Luo, Bobo Li, Shanqing Xu, Shize Zhang et al.ICLR 2026 · 10 citations
- MetaMind: Modeling Human Social Thoughts with Metacognitive Multi-Agent SystemsXuanming Zhang, Yuxuan Chen, Samuel (Min-Hsuan) Yeh, Sharon LiNeurIPS 2025 · 14 citations
- EmotionThinker: Prosody-Aware Reinforcement Learning for Explainable Speech Emotion ReasoningDingdong WANG, Shujie LIU, Tianhua Zhang, Youjun Chen et al.ICLR 2026 · 22 citations
- Modeling Protagonist Emotions for Emotion-Aware StorytellingFaeze Brahman, Snigdha ChaturvediEMNLP 2020 · 5 citations
