Emotions Where Art Thou: Understanding and Characterizing the Emotional Latent Space of Large Language Models
Benjamin Z. Reichman, Adar Avsian, Larry Heck
摘要
This work investigates how large language models (LLMs) internally represent emotion by analyzing the geometry of their hidden-state space. The paper identifies a low-dimensional emotional manifold and shows that emotional representations are directionally encoded, distributed across layers, and aligned with interpretable dimensions. These structures are stable across depth and generalize to eight realworld emotion datasets spanning five languages. Cross-domain alignment yields low error and strong linear probe performance, indicating a universal emotional subspace. Within this space, internal emotion perception can be steered while preserving semantics using a learned intervention module, with especially strong control for basic emotions across languages. These findings reveal a consistent and manipulable affective geometry in LLMs and offer insight into how they internalize and process emotion. Recent work has also examined emotion manipulation and decoding. For instance, models have been used to map text to dimensional emotion ratings like valence-arousal-dominance (VAD) (Shah et al., 2023; Broekens et al., 2023) , or to generate emotionally inflected language on demand (Reichman et al., 2025) . LLMs have also been shown to be more likely to comply with emotionally framed requests (Vinay et al., 2024) . These studies also treat emotion primarily as a label or generation condition-not a latent internal representation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Mapping the Circumplex of Affect: Geometric Analysis of Emotion Representations via Hyperspherical Contrastive LearningYusuke Yamauchi, Akiko AizawaACL 2026
- CoCoEmo: Composable and Controllable Human-Like Emotional TTS via Activation SteeringSiyi Wang, Shihong Tan, Siyi Liu, Hong Jia 等ICML 2026
它引用的顶会 Paper10
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Plug and Play Language Models: A Simple Approach to Controlled Text GenerationSumanth Dathathri, Andrea Madotto, Janice Lan, Jane Hung 等ICLR 2020 · 被引用 1,166 次
- Apathetic or Empathetic? Evaluating LLMs' Emotional Alignments with HumansJen-tse Huang, Man Ho Lam, Eric John Li, Shujie Ren 等NeurIPS 2024 · 被引用 63 次
- Whispering Experts: Neural Interventions for Toxicity Mitigation in Language ModelsXavier Suau, Pieter Delobelle, Katherine Metcalf, Armand Joulin 等ICML 2024 · 被引用 31 次
- GoEmotions: A Dataset of Fine-Grained EmotionsDorottya Demszky, Dana Movshovitz-Attias, Jeongwoo Ko, Alan S. Cowen 等ACL 2020 · 被引用 16 次
相关 Paper
- Do LLMs “Feel”? Emotion Circuits Discovery and ControlChenxi Wang, Yixuan Zhang, Ruiji Yu, Yufei Zheng 等ICML 2026 · 被引用 11 次
- Exploring Modular Prompt Design for Emotion and Mental Health RecognitionMinseo Kim, Taemin Kim, Thu Hoang Anh Vo, Yugyeong Jung 等CHI 2025 · 被引用 11 次
- Sparse Autoencoders for Interpretable Emotion Control in Text-to-SpeechHongfei Du, Jiacheng Shi, Sidi Lu, Gang Zhou 等ICML 2026
- Characterizing and Evaluating Working Emotion Vocabularies in Multilingual Large Language ModelsNicholas Deas, Iván Ernesto Pérez Mejía, Ellie Yang, Kathleen McKeownACL 2026
- Beyond Single Concept Vector: Modeling Concept Subspace in LLMs with Gaussian DistributionHaiyan Zhao, Heng Zhao, Bo Shen, Ali Payani 等ICLR 2025 · 被引用 2 次
