Social Agents: Collective Intelligence Improves LLM Predictions
Aanisha Bhattacharyya, Abhilekh Borah, Yaman Singla, Rajiv Ratn Shah, Changyou Chen, Balaji Krishnamurthy
Abstract
In human society, collective decision making has often outperformed the judgment of individuals. Classic examples range from estimating livestock weights to predicting elections and financial markets, where averaging many independent guesses often yields results more accurate than those of experts. These successes arise because groups bring together diverse perspectives, independent voices, and distributed knowledge, so that idiosyncratic errors average out across independent estimates rather than compound. This principle, known as the Wisdom of Crowds, underpins forecasting in domains from finance to politics. Large Language Models (LLMs), however, typically produce a single definitive answer. While effective in many settings, this uniformity overlooks the diversity of human judgments that shapes how people respond to ads, videos, and webpages. Inspired by how societies benefit from diverse opinions, we ask whether LLM predictions can be improved by simulating many diverse answers rather than one. We introduce Social Agents, a multi-agent framework that instantiates a synthetic society of human-like personas with diverse demographic (e.g., age, gender) and psychographic (e.g., values, interests) attributes. Each persona independently appraises a stimulus such as an advertisement, video, or webpage, offering both a quantitative score (e.g., click-through likelihood, recall score, likability) and a qualitative rationale. The set of persona opinions mirrors a real human crowd, and aggregating them yields a single estimate closer to the crowd mean than any individual estimate. Across eleven behavioral prediction tasks, Social Agents outperforms single-LLM baselines by up to 164% on simple judgments (e.g., webpage likability) and up to 24% on complex interpretive reasoning (e.g., video memorability), both with GPT-4o as the backbone. Averaged across models, gains reach 30.5% on low-level and 9.9% on high-level tasks. The individual persona predictions generated by Social Agents also strongly align with human judgments, reaching Pearson correlations up to 0.71. These results position computational crowd simulation as a scalable, interpretable tool for improving behavioral prediction and supporting behavioral and marketing decisions
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2c59928f-7511-442a-837b-2b8153c3cfc1Builds on7
- Whose Opinions Do Language Models Reflect?Shibani Santurkar, Esin Durmus, Faisal Ladhak, Cinoo Lee et al.ICML 2023 · 764 citations
- Self-Consistency Improves Chain of Thought Reasoning in Language ModelsXuezhi Wang, Jason Wei, Dale Schuurmans, Quoc V. Le et al.ICLR 2023 · 681 citations
- Quantifying the Persona Effect in LLM SimulationsTiancheng Hu, Nigel CollierACL 2024 · 22 citations
- Documenting Large Webtext Corpora: A Case Study on the Colossal Clean Crawled CorpusJesse Dodge, Maarten Sap, Ana Marasovic, William Agnew et al.EMNLP 2021 · 18 citations
- GoEmotions: A Dataset of Fine-Grained EmotionsDorottya Demszky, Dana Movshovitz-Attias, Jeongwoo Ko, Alan S. Cowen et al.ACL 2020 · 16 citations
Related papers
- MetaAgents: Large Language Model Based Agents for Decision-Making on TeamingYuan Li, Lichao Sun, Yixuan ZhangCSCW 2025 · 35 citations
- Social Dynamics as Critical Vulnerabilities that Undermine Objective Decision-Making in LLM CollectivesChanggeon Ko, Jisu Shin, Hoyun Song, Huije Lee et al.ACL 2026 · 1 citation
- To Mask or to Mirror: Human-AI Alignment in Collective ReasoningCrystal Qian, Aaron T. Parisi, Clémentine Bouleau, Vivian Tsai et al.EMNLP 2025 · 1 citation
- AgentVerse: Facilitating Multi-Agent Collaboration and Exploring Emergent BehaviorsWeize Chen, Yusheng Su, Jingwei Zuo, Cheng Yang et al.ICLR 2024 · 594 citations
- Can Large Language Model Agents Simulate Human Trust Behavior?Chengxing Xie, Canyu Chen, Feiran Jia, Ziyu Ye et al.NeurIPS 2024 · 183 citations
