LLM vs Small Model? Large Language Model Based Text Augmentation Enhanced Personality Detection Model
Linmei Hu, Hongyu He, Duokang Wang, Ziwang Zhao, Yingxia Shao, Liqiang Nie
摘要
Personality detection aims to detect one's personality traits underlying in social media posts. One challenge of this task is the scarcity of ground-truth personality traits which are collected from self-report questionnaires. Most existing methods learn post features directly by fine-tuning the pre-trained language models under the supervision of limited personality labels. This leads to inferior quality of post features and consequently affects the performance. In addition, they treat personality traits as one-hot classification labels, overlooking the semantic information within them. In this paper, we propose a large language model (LLM) based text augmentation enhanced personality detection model, which distills the LLM's knowledge to enhance the small model for personality detection, even when the LLM fails in this task. Specifically, we enable LLM to generate post analyses (augmentations) from the aspects of semantic, sentiment, and linguistic, which are critical for personality detection. By using contrastive learning to pull them together in the embedding space, the post encoder can better capture the psycho-linguistic information within the post representations, thus improving personality detection. Furthermore, we utilize the LLM to enrich the information of personality labels for enhancing the detection performance. Experimental results on the benchmark datasets demonstrate that our model outperforms the state-of-the-art methods on personality detection.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- EmoMAS: Emotion-Aware Multi-Agent System for High-Stakes Edge-Deployable Negotiation with Bayesian OrchestrationYunbo Long, Yuhan Liu, Liming XuACL 2026 · 被引用 6 次
- Leveraging the Dual Capabilities of LLM: LLM-Enhanced Text Mapping Model for Personality DetectionWeihong Bi, Feifei Kou, Lei Shi, Yawen Li 等AAAI 2025 · 被引用 3 次
- Persona-E²: A Human-Grounded Dataset for Personality-Shaped Emotional Responses to Textual EventsYuqin Yang, Haowu Zhou, Haoran Tu, Zhiwen Hui 等ACL 2026
- LLM-Driven Completeness and Consistency Evaluation for Cultural Heritage Data Augmentation in Cross-Modal RetrievalJian Zhang, Junyi Guo, Junyi Yuan, Huanda Lu 等EMNLP 2025
- Towards Transferable Personality Representation Learning based on Triplet Comparisons and Its ApplicationsKai Tang, Rui Wang, Renyu Zhu, Minmin Lin 等EMNLP 2025
它引用的顶会 Paper9
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 被引用 2,496 次
- Self-Instruct: Aligning Language Models with Self-Generated InstructionsYizhong Wang, Yeganeh Kordi, Swaroop Mishra, Alisa Liu 等ACL 2023 · 被引用 540 次
- Hierarchical Modeling for User Personality Prediction: The Role of Message-Level AttentionVeronica E. Lynn, Niranjan Balasubramanian, H. Andrew SchwartzACL 2020 · 被引用 72 次
相关 Paper
- Data Augmented Graph Neural Networks for Personality DetectionYangfu Zhu, Yue Xia, Meiling Li, Tingting Zhang 等AAAI 2024 · 被引用 15 次
- Knowledge-Enhanced Hierarchical Heterogeneous Graph for Personality Identification with Limited Training DataYuxuan Song, Qiudan Li, Yilin Wu, David Jingjun Xu 等AAAI 2025
- Multi-Document Transformer for Personality DetectionFeifan Yang, Xiaojun Quan, Yunyi Yang, Jianxing YuAAAI 2021 · 被引用 57 次
- Towards S²-Challenges Underlying LLM-Based Augmentation for Personalized News RecommendationShicheng Wang, Hengzhu Tang, Li Gao, Shu Guo 等AAAI 2025 · 被引用 2 次
- Miko: Multimodal Intention Knowledge Distillation from Large Language Models for Social-Media Commonsense DiscoveryFeihong Lu, Weiqi Wang, Yangyifei Luo, Ziqin Zhu 等ACM MM 2024 · 被引用 11 次
