Alignment with human representations supports robust few-shot learning
Ilia Sucholutsky, Tom Griffiths
Abstract
Should we care whether AI systems have representations of the world that are similar to those of humans? We provide an information-theoretic analysis that suggests that there should be a U-shaped relationship between the degree of representational alignment with humans and performance on few-shot learning tasks. We confirm this prediction empirically, finding such a relationship in an analysis of the performance of 491 computer vision models. We also show that highly-aligned models are more robust to both natural adversarial attacks and domain shifts. Our results suggest that human-alignment is often a sufficient, but not necessary, condition for models to make effective use of limited data, be robust, and generalize well.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 05f2a044-5aa0-465d-acc6-492d959b81e1Cited by top-tier papers6
- Improving neural network representations using human similarity judgmentsLukas Muttenthaler, Lorenz Linhardt, Jonas Dippel, Robert A. Vandermeulen et al.NeurIPS 2023 · 61 citations
- Performance-optimized deep neural networks are evolving into worse models of inferotemporal visual cortexDrew Linsley, Ivan F. Rodriguez Rodriguez, Thomas Fel, Michael Arcaro et al.NeurIPS 2023 · 38 citations
- What Matters to You? Towards Visual Representation Alignment for Robot LearningThomas Tian, Chenfeng Xu, Masayoshi Tomizuka, Jitendra Malik et al.ICLR 2024 · 17 citations
- Evaluating alignment between humans and neural network representations in image-based learning tasksCan Demircan, Tankred Saanum, Leonardo Pettini, Marcel Binz et al.NeurIPS 2024 · 11 citations
- Learning Human-like Representations to Enable Learning Human ValuesAndrea Wynn, Ilia Sucholutsky, Tom GriffithsNeurIPS 2024 · 11 citations
Builds on10
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- BEiT: BERT Pre-Training of Image TransformersHangbo Bao, Li Dong, Songhao Piao, Furu WeiICLR 2022 · 3,632 citations
- The Many Faces of Robustness: A Critical Analysis of Out-of-Distribution GeneralizationDan Hendrycks, Steven Basart, Norman Mu, Saurav Kadavath et al.ICCV 2021 · 2,294 citations
- Fine-Tuning can Distort Pretrained Features and Underperform Out-of-DistributionAnanya Kumar, Aditi Raghunathan, Robbie Matthew Jones, Tengyu Ma et al.ICLR 2022 · 911 citations
- High-Performance Large-Scale Image Recognition Without NormalizationAndy Brock, Soham De, Samuel L. Smith, Karen SimonyanICML 2021 · 613 citations
Related papers
- Human alignment of neural network representationsLukas Muttenthaler, Jonas Dippel, Lorenz Linhardt, Robert A. Vandermeulen et al.ICLR 2023 · 15 citations
- Channel Importance Matters in Few-Shot Image ClassificationXu Luo, Jing Xu, Zenglin XuICML 2022 · 57 citations
- Few-Shot Adversarial Prompt Learning on Vision-Language ModelsYiwei Zhou, Xiaobo Xia, Zhiwei Lin, Bo Han et al.NeurIPS 2024 · 48 citations
- Simple Semantic-Aided Few-Shot LearningHai Zhang, Junzhe Xu, Shanlin Jiang, Zhenan HeCVPR 2024 · 33 citations
- BIRD: Behavior Induction via Representation-structure DistillationGalen Pogoncheff, Michael BeyelerICLR 2026
