LLM vs Small Model? Large Language Model Based Text Augmentation Enhanced Personality Detection Model
Linmei Hu, Hongyu He, Duokang Wang, Ziwang Zhao, Yingxia Shao, Liqiang Nie
Abstract
Personality detection aims to detect one's personality traits underlying in social media posts. One challenge of this task is the scarcity of ground-truth personality traits which are collected from self-report questionnaires. Most existing methods learn post features directly by fine-tuning the pre-trained language models under the supervision of limited personality labels. This leads to inferior quality of post features and consequently affects the performance. In addition, they treat personality traits as one-hot classification labels, overlooking the semantic information within them. In this paper, we propose a large language model (LLM) based text augmentation enhanced personality detection model, which distills the LLM's knowledge to enhance the small model for personality detection, even when the LLM fails in this task. Specifically, we enable LLM to generate post analyses (augmentations) from the aspects of semantic, sentiment, and linguistic, which are critical for personality detection. By using contrastive learning to pull them together in the embedding space, the post encoder can better capture the psycho-linguistic information within the post representations, thus improving personality detection. Furthermore, we utilize the LLM to enrich the information of personality labels for enhancing the detection performance. Experimental results on the benchmark datasets demonstrate that our model outperforms the state-of-the-art methods on personality detection.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 55516c83-1ab9-4d4b-9940-40664240e3faCited by top-tier papers7
- EmoMAS: Emotion-Aware Multi-Agent System for High-Stakes Edge-Deployable Negotiation with Bayesian OrchestrationYunbo Long, Yuhan Liu, Liming XuACL 2026 · 6 citations
- Leveraging the Dual Capabilities of LLM: LLM-Enhanced Text Mapping Model for Personality DetectionWeihong Bi, Feifei Kou, Lei Shi, Yawen Li et al.AAAI 2025 · 3 citations
- Persona-E²: A Human-Grounded Dataset for Personality-Shaped Emotional Responses to Textual EventsYuqin Yang, Haowu Zhou, Haoran Tu, Zhiwen Hui et al.ACL 2026
- LLM-Driven Completeness and Consistency Evaluation for Cultural Heritage Data Augmentation in Cross-Modal RetrievalJian Zhang, Junyi Guo, Junyi Yuan, Huanda Lu et al.EMNLP 2025
- Towards Transferable Personality Representation Learning based on Triplet Comparisons and Its ApplicationsKai Tang, Rui Wang, Renyu Zhu, Minmin Lin et al.EMNLP 2025
Builds on9
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 2,496 citations
- Self-Instruct: Aligning Language Models with Self-Generated InstructionsYizhong Wang, Yeganeh Kordi, Swaroop Mishra, Alisa Liu et al.ACL 2023 · 540 citations
- Hierarchical Modeling for User Personality Prediction: The Role of Message-Level AttentionVeronica E. Lynn, Niranjan Balasubramanian, H. Andrew SchwartzACL 2020 · 72 citations
Related papers
- Data Augmented Graph Neural Networks for Personality DetectionYangfu Zhu, Yue Xia, Meiling Li, Tingting Zhang et al.AAAI 2024 · 15 citations
- Knowledge-Enhanced Hierarchical Heterogeneous Graph for Personality Identification with Limited Training DataYuxuan Song, Qiudan Li, Yilin Wu, David Jingjun Xu et al.AAAI 2025
- Multi-Document Transformer for Personality DetectionFeifan Yang, Xiaojun Quan, Yunyi Yang, Jianxing YuAAAI 2021 · 57 citations
- Towards S²-Challenges Underlying LLM-Based Augmentation for Personalized News RecommendationShicheng Wang, Hengzhu Tang, Li Gao, Shu Guo et al.AAAI 2025 · 2 citations
- Miko: Multimodal Intention Knowledge Distillation from Large Language Models for Social-Media Commonsense DiscoveryFeihong Lu, Weiqi Wang, Yangyifei Luo, Ziqin Zhu et al.ACM MM 2024 · 11 citations
