Brain Decodes Deep Nets
Huzheng Yang, James C. Gee, Jianbo Shi
摘要
We developed a tool for visualizing and analyzing large pre-trained vision models by mapping them onto the brain, thus exposing their hidden inside. Our innovation arises from a surprising usage of brain encoding: predicting brain fMRI measurements in response to images. We report two findings. First, explicit mapping between the brain and deep-network features across dimensions of space, layers, scales, and channels is crucial. This mapping method, Fac-torTopy, is plug-and-play for any deep-network; with it, one can paint a picture of the network onto the brain (liter-ally!). Second, our visualization shows how different training methods matter: they lead to remarkable differences in hierarchical organization and scaling behavior, growing with more data or network capacity. It also provides in-sight into fine-tuning: how pre-trained models change when adapting to small datasets. We found brain-like hierarchi-cally organized network suffer less from catastrophic for-getting after fine-tuned.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- Transformer brain encoders explain human high-level visual responsesHossein Adeli, Minni Sun, Nikolaus KriegeskorteNeurIPS 2025 · 被引用 14 次
- In Silico Mapping of Visual Categorical Selectivity Across the Whole BrainEthan Hwang, Hossein Adeli, Wenxuan Guo, Andrew F. Luo 等NeurIPS 2025 · 被引用 7 次
- Meta-Learning an In-Context Transformer Model of Human Higher Visual CortexMuquan Yu, Mu Nan, Hossein Adeli, Jacob S. Prince 等NeurIPS 2025 · 被引用 5 次
- A Cognitive Process-Inspired Architecture for Subject-Agnostic Brain Visual DecodingJingyu Lu, Haonan Wang, Qixiang Zhang, Xiaomeng LiICLR 2026 · 被引用 3 次
- HyFI: Hyperbolic Feature Interpolation for Brain-Vision AlignmentSangmin Jo, Wootaek Jeong, Da-Woon Heo, Yoohwan Hwang 等AAAI 2026 · 被引用 2 次
它引用的顶会 Paper17
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- An Empirical Study of Training Self-Supervised Vision TransformersXinlei Chen, Saining Xie, Kaiming HeICCV 2021 · 被引用 2,340 次
- A Tale of Two Features: Stable Diffusion Complements DINO for Zero-Shot Semantic CorrespondenceJunyi Zhang, Charles Herrmann, Junhwa Hur, Luisa Polania Cabrera 等NeurIPS 2023 · 被引用 371 次
相关 Paper
- Disentangling the Factors of Convergence between Brains and DINOv3Joséphine Raugel, Marc Szafraniec, Huy V. Vo, Camille Couprie 等ICLR 2026
- One Hundred Neural Networks and Brains Watching Videos: Lessons from AlignmentChristina Sartzetaki, Gemma Roig, Cees G. M. Snoek, Iris I. A. GroenICLR 2025
- Dimensionality Mismatch Between Brains and Artificial Neural NetworksSantiago Galella, Maren H. Wehrheim, Matthias KaschubeNeurIPS 2025
- Multimodal Scaling Laws for Task & Data-Optimized Models of Visual CortexAbdülkadir Gökce, Yingtian Tang, Martin SchrimpfICML 2026
- Brain encoding models based on multimodal transformers can transfer across language and visionJerry Tang, Meng Du, Vy A. Vo, Vasudev Lal 等NeurIPS 2023 · 被引用 76 次
