When Language Models Lose Their Mind: The Consequences of Brain Misalignment
Gabriele Merlin, Mariya Toneva
Abstract
While brain-aligned large language models (LLMs) have garnered attention for their potential as cognitive models and for potential for enhanced safety and trustworthiness in AI, the role of this brain alignment for linguistic competence remains uncertain. In this work, we investigate the functional implications of brain alignment by introducing brain-misaligned models--LLMs intentionally trained to predict brain activity poorly while maintaining high language modeling performance. We evaluate these models on over 200 downstream tasks encompassing diverse linguistic domains, including semantics, syntax, discourse, reasoning, and morphology. By comparing brain-misaligned models with well-matched brain-aligned counterparts, we isolate the specific impact of brain alignment on language understanding. Our experiments reveal that brain misalignment substantially impairs downstream performance, highlighting the critical role of brain alignment in achieving robust linguistic competence. These findings underscore the importance of brain alignment in LLMs and offer novel insights into the relationship between neural representations and linguistic processing.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b5188da5-cd91-4fa2-9e9e-c48d6366affaBuilds on10
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Joint processing of linguistic properties in brains and language modelsSubba Reddy Oota, Manish Gupta, Mariya TonevaNeurIPS 2023 · 64 citations
- Can fMRI reveal the representation of syntactic structure in the brain?Aniketh Janardhan Reddy, Leila WehbeNeurIPS 2021 · 54 citations
- Prompting Language Models for Linguistic StructureTerra Blevins, Hila Gonen, Luke ZettlemoyerACL 2023 · 15 citations
- Training language models to summarize narratives improves brain alignmentKhai Loong Aw, Mariya TonevaICLR 2023 · 11 citations
Related papers
- From Language to Cognition: How LLMs Outgrow the Human Language NetworkBadr AlKhamissi, Greta Tuckute, Yingtian Tang, Taha Osama A Binhuraib et al.EMNLP 2025 · 1 citation
- Language models and brains align due to more than next-word prediction and word-level informationGabriele Merlin, Mariya TonevaEMNLP 2024 · 2 citations
- Linguistic Properties and Model Scale in Brain Encoding: From Small to Compressed Language ModelsSubba Reddy Oota, Satya Sai Srinath Namburi GNVV, Vijay Rowtula, Khushbu Pahwa et al.ICML 2026
- Do Large Language Models Think like the Brain? Sentence-Level Evidences from Layer-Wise Embeddings and fMRIYu Lei, Xingyang Ge, Yi Zhang, Yiming Yang et al.AAAI 2026 · 2 citations
- Fine-grained Analysis of Brain-LLM Alignment through Input AttributionMichela Proietti, Roberto Capobianco, Mariya TonevaICML 2026
