Evolution and compression in LLMs: on the emergence of human-aligned categorization
Nathaniel Imel, Noga Zaslavsky
摘要
Converging evidence suggests that human systems of semantic categories achieve near-optimal compression via the Information Bottleneck (IB) complexity-accuracy tradeoff. Large language models (LLMs) are not trained for this objective, which raises the question: are LLMs capable of evolving efficient human-aligned semantic systems? To address this question, we focus on color categorization --- a key testbed of cognitive theories of categorization with uniquely rich human data --- and replicate with LLMs two influential human studies. First, we conduct an English color-naming study, showing that LLMs vary widely in their complexity and English-alignment, with larger instruction-tuned models achieving better alignment and IB-efficiency. Second, to test whether these LLMs simply mimic patterns in their training data or actually exhibit a human-like inductive bias toward IB-efficiency, we simulate cultural evolution of pseudo color-naming systems in LLMs via a method we refer to as Iterated in-Context Language Learning (IICLL). We find that akin to humans, LLMs iteratively restructure initially random systems towards greater IB-efficiency. However, only a model with strongest in-context capabilities (Gemini 2.0) is able to recapitulate the wide range of near-optimal IB-tradeoffs observed in humans, while other state-of-the-art models converge to low-complexity solutions. These findings demonstrate how human-aligned semantic categories can emerge in LLMs via the same fundamental principle that underlies semantic efficiency in humans.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- Compositional languages emerge in a neural iterated learning modelYi Ren, Shangmin Guo, Matthieu Labeau, Shay B. Cohen 等ICLR 2020 · 被引用 111 次
- Emergent Communication at ScaleRahma Chaabouni, Florian Strub, Florent Altché, Eugene Tarassov 等ICLR 2022 · 被引用 65 次
- Trading off Utility, Informativeness, and Complexity in Emergent CommunicationMycal Tucker, Roger Levy, Julie A. Shah, Noga ZaslavskyNeurIPS 2022 · 被引用 34 次
- Bias Amplification in Language Model Evolution: An Iterated Learning PerspectiveYi Ren, Shangmin Guo, Linlu Qiu, Bailin Wang 等NeurIPS 2024 · 被引用 23 次
- Bridging semantics and pragmatics in information-theoretic emergent communicationEleonora Gualdoni, Mycal Tucker, Roger Levy, Noga ZaslavskyNeurIPS 2024 · 被引用 7 次
相关 Paper
- From Tokens to Thoughts: How LLMs and Humans Trade Compression for MeaningChen Shani, Liron Soffer, Dan Jurafsky, Yann LeCun 等ICLR 2026 · 被引用 38 次
- Learning is Forgetting; LLM Training As Lossy CompressionHenry Conklin, Tom Hosking, Yi Chern Tan, Jonathan D. Cohen 等ICLR 2026 · 被引用 6 次
- To Mask or to Mirror: Human-AI Alignment in Collective ReasoningCrystal Qian, Aaron T. Parisi, Clémentine Bouleau, Vivian Tsai 等EMNLP 2025 · 被引用 1 次
- Limits of Transformer Language Models on Learning to Compose AlgorithmsJonathan Thomm, Giacomo Camposampiero, Aleksandar Terzic, Michael Hersche 等NeurIPS 2024 · 被引用 16 次
- Measuring Intent Comprehension in LLMsNadav Kunievsky, James EvansICML 2026 · 被引用 1 次
