Different types of syntactic agreement recruit the same units within large language models
Daria Kryvosheieva, Andrea Gregor de Varda, Evelina Fedorenko, Greta Tuckute
Abstract
Large language models (LLMs) can reliably distinguish grammatical from ungrammatical sentences, but how grammatical knowledge is represented within the models remains an open question. We investigate whether different syntactic phenomena recruit shared or distinct components in LLMs. Using a functional localization approach inspired by cognitive neuroscience, we identify the LLM units most responsive to 67 English syntactic phenomena in seven open-weight models. These units are consistently recruited across sentences containing the phenomena and causally support the models' syntactic performance. Critically, different types of syntactic agreement (e.g., subject-verb, anaphor, determiner-noun) recruit overlapping sets of units, suggesting that agreement constitutes a meaningful functional category for LLMs. This pattern holds in English, Russian, and Chinese; and further, in a cross-lingual analysis of 57 diverse languages, structurally more similar languages share more units for subject-verb agreement. Taken together, these findings reveal that syntactic agreement-a critical marker of syntactic dependencies-constitutes a meaningful category within LLMs' representational spaces. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 57a313b2-0dd2-4701-92ee-2589d313f020Cited by top-tier papers1
Ask how each one uses itBuilds on17
- Interpretability at Scale: Identifying Causal Mechanisms in AlpacaZhengxuan Wu, Atticus Geiger, Thomas Icard, Christopher Potts et al.NeurIPS 2023 · 146 citations
- Can fMRI reveal the representation of syntactic structure in the brain?Aniketh Janardhan Reddy, Leila WehbeNeurIPS 2021 · 54 citations
- Understanding and Minimising Outlier Features in Transformer TrainingBobby He, Lorenzo Noci, Daniele Paliotta, Imanol Schlag et al.NeurIPS 2024 · 27 citations
- Probing Word Syntactic Representations in the Brain by a Feature Elimination MethodXiaohan Zhang, Shaonan Wang, Nan Lin, Jiajun Zhang et al.AAAI 2022 · 26 citations
- Paths Not Taken: Understanding and Mending the Multilingual Factual Recall PipelineMeng Lu, Ruochen Zhang, Carsten Eickhoff, Ellie PavlickEMNLP 2025 · 15 citations
Related papers
- Causal Analysis of Syntactic Agreement Mechanisms in Neural Language ModelsMatthew Finlayson, Aaron Mueller, Sebastian Gehrmann, Stuart M. Shieber et al.ACL 2021
- Word Frequency Does Not Predict Grammatical Knowledge in Language ModelsCharles Yu, Ryan Sie, Nico Tedeschi, Leon BergenEMNLP 2020 · 6 citations
- Structural Priming Demonstrates Abstract Grammatical Representations in Multilingual Language ModelsJames A. Michaelov, Catherine Arnett, Tyler A. Chang, Ben BergenEMNLP 2023 · 6 citations
- Syntactic Substitutability as Unsupervised Dependency SyntaxJasper Jian, Siva ReddyEMNLP 2023
- Controlled Evaluation of Grammatical Knowledge in Mandarin Chinese Language ModelsYiwen Wang, Jennifer Hu, Roger Levy, Peng QianEMNLP 2021 · 3 citations
