On the Reliability of Psychological Scales on Large Language Models
Jen-tse Huang, Wenxiang Jiao, Man Ho Lam, Eric John Li, Wenxuan Wang, Michael R. Lyu
Abstract
Recent research has focused on examining Large Language Models' (LLMs) characteristics from a psychological standpoint, acknowledging the necessity of understanding their behavioral characteristics.The administration of personality tests to LLMs has emerged as a noteworthy area in this context.However, the suitability of employing psychological scales, initially devised for humans, on LLMs is a matter of ongoing debate.Our study aims to determine the reliability of applying personality assessments to LLMs, explicitly investigating whether LLMs demonstrate consistent personality traits.Analysis of 2,500 settings per model, including GPT-3.5, GPT-4, Gemini-Pro, and LLaMA-3.1, reveals that various LLMs show consistency in responses to the Big Five Inventory, indicating a satisfactory level of reliability.Furthermore, our research explores the potential of GPT-3.5 to emulate diverse personalities and represent various groups-a capability increasingly sought after in social sciences for substituting human participants with LLMs to reduce costs.Our findings reveal that LLMs have the potential to represent different personalities with specific prompt instructions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 89de4367-1233-4b6e-8745-c7d0e77667fbCited by top-tier papers8
- On the Humanity of Conversational AI: Evaluating the Psychological Portrayal of LLMsJen-tse Huang, Wenxuan Wang, Eric John Li, Man Ho Lam et al.ICLR 2024 · 85 citations
- Apathetic or Empathetic? Evaluating LLMs' Emotional Alignments with HumansJen-tse Huang, Man Ho Lam, Eric John Li, Shujie Ren et al.NeurIPS 2024 · 63 citations
- AI-exhibited Personality Traits Can Shape Human Self-concept through ConversationsJingshu Li, Tianqi Song, Nattapat Boonprakong, Zicheng Zhu et al.CHI 2026 · 3 citations
- Interview-Informed Generative Agents for Product Discovery: A Validation StudyZichao Wang, Alexa F. SiuCHI 2026 · 1 citation
- Quantifying and Mitigating Socially Desirable Responding in LLMs: A Desirability-Matched Graded Forced-Choice Psychometric StudyKensuke Okada, Yui Furukawa, Kyosuke BunjiACL 2026 · 1 citation
Builds on9
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- Calibrate Before Use: Improving Few-shot Performance of Language ModelsZihao Zhao, Eric Wallace, Shi Feng, Dan Klein et al.ICML 2021 · 1,843 citations
- Whose Opinions Do Language Models Reflect?Shibani Santurkar, Esin Durmus, Faisal Ladhak, Cinoo Lee et al.ICML 2023 · 764 citations
- Large Language Models Are Zero-Shot Fuzzers: Fuzzing Deep-Learning Libraries via Large Language ModelsYinlin Deng, Chunqiu Steven Xia, Haoran Peng, Chenyuan Yang et al.ISSTA 2023 · 253 citations
- Automated Repair of Programs from Large Language ModelsZhiyu Fan, Xiang Gao, Martin Mirchev, Abhik Roychoudhury et al.ICSE 2023 · 213 citations
Related papers
- Open Models, Closed Minds? On Agents Capabilities in Mimicking Human Personalities through Open Large Language ModelsLucio La Cava, Andrea TagarelliAAAI 2025 · 36 citations
- Can LLM Agents Maintain a Persona in Discourse?Pranav Bhandari, Nicolas Fay, Michael J. Wise, Amitava Datta et al.EMNLP 2025
- Evaluating Psychological Safety of Large Language ModelsXingxuan Li, Yutong Li, Lin Qiu, Shafiq Joty et al.EMNLP 2024 · 14 citations
- The Personality Dimensions GPT-3 Expresses During Human-Chatbot InteractionsNikola Kovacevic, Christian Holz, Markus Gross, Rafael WampflerUbiComp 2024 · 15 citations
- To Mask or to Mirror: Human-AI Alignment in Collective ReasoningCrystal Qian, Aaron T. Parisi, Clémentine Bouleau, Vivian Tsai et al.EMNLP 2025 · 1 citation
