Inducing Dyslexia in Vision Language Models
Melika Honarmand, Ayati Sharma, Badr AlKhamissi, Johannes Mehrer, Martin Schrimpf
摘要
Dyslexia, a neurodevelopmental disorder characterized by persistent reading difficulties, is often linked to reduced activity of the visual word form area (VWFA) in the ventral occipito-temporal cortex. Traditional approaches to studying dyslexia, such as behavioral and neuroimaging methods, have provided valuable insights but remain limited in their ability to test causal hypotheses about the underlying mechanisms of reading impairments. In this study, we use large-scale vision-language models (VLMs) to simulate dyslexia by functionally identifying and perturbing artificial analogues of word processing. Using stimuli from cognitive neuroscience, we identify visual-word-form-selective units within VLMs and demonstrate that they predict human VWFA neural responses. Ablating model VWF units leads to selective impairments in reading tasks while general visual and language comprehension abilities remain intact. In particular, the resulting model matches dyslexic humans' phonological deficits without a significant change in orthographic processing, and mirrors dyslexic behavior in font sensitivity. Taken together, our modeling results replicate key characteristics of dyslexia and establish a computational framework for investigating brain disorders. 1 * Joint supervision 1 Code available via GitHub.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- Scaling Laws for Task-Optimized Models of the Primate Visual Ventral StreamAbdülkadir Gökce, Martin SchrimpfICML 2025
- TopoLM: brain-like spatio-functional organization in a topographic language modelNeil Rathi, Johannes Mehrer, Badr AlKhamissi, Taha Osama A Binhuraib 等ICLR 2025
- Contour Integration Underlies Human-Like VisionBen Lonnqvist, Elsa Scialom, Abdulkadir Gokce, Zehra Merchant 等ICML 2025
相关 Paper
- Speech language models lack important brain-relevant semanticsSubba Reddy Oota, Emin Çelik, Fatma Deniz, Mariya TonevaACL 2024
- Grounding Visual Illusions in Language: Do Vision-Language Models Perceive Illusions Like Humans?Yichi Zhang, Jiayi Pan, Yuchen Zhou, Rui Pan 等EMNLP 2023 · 被引用 9 次
- With Ears to See and Eyes to Hear: Sound Symbolism Experiments with Multimodal Large Language ModelsTyler Loakman, Yucheng Li, Chenghua LinEMNLP 2024 · 被引用 1 次
- Do VLMs Perceive or Recall? Probing Visual Perception vs. Memory with Classic Visual IllusionsXiaoxiao Sun, Mingyang Li, Kun Yuan, Min Woo Sun 等CVPR 2026 · 被引用 8 次
- DYPA: A Machine Learning Dyslexia Prescreening Mobile Application for Chinese ChildrenShuhan Zhong, Sizhe Song, Tianhao Tang, Fei Nie 等UbiComp 2023 · 被引用 9 次
