Can Biases in ImageNet Models Explain Generalization?
Paul Gavrikov, Janis Keuper
Abstract
The robust generalization of models to rare, indistribution (ID) samples drawn from the long tail of the training distribution and to out-of-training-distribution (OOD) samples is one of the major challenges of current deep learning methods. For image classification, this manifests in the existence of adversarial attacks, the performance drops on distorted images, and a lack of generalization to concepts such as sketches. The current understanding of generalization in neural networks is very limited, but some biases that differentiate models from human vision have been identified and might be causing these limitations. Consequently, several attempts with varying success have been made to reduce these biases during training to improve generalization. We take a step back and sanitycheck these attempts. Fixing the architecture to the wellestablished ResNet-50, we perform a large-scale study on 48 ImageNet models obtained via different training methods to understand how and if these biases -including shape bias, spectral biases, and critical bands -interact with generalization. Our extensive study results reveal that contrary to previous findings, these biases are insufficient to accurately predict the generalization of a model holistically. We provide access to all checkpoints and evaluation code at https://github.com/paulgavrikov/biases_ vs_generalization/
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers12
- ImageNet-trained CNNs are not biased towards texture: Revisiting feature reliance through controlled suppressionTom Burgert, Oliver Stoll, Paolo Rota, Begüm DemirNeurIPS 2025 · 19 citations
- TinyViM: Frequency Decoupling for Tiny Hybrid Vision MambaXiaowen Ma, Zhenliang Ni, Xinghao ChenICCV 2025 · 16 citations
- When Pretty Isn't Useful: Investigating Why Modern Text-to-Image Models Fail as Reliable Training Data GeneratorsKrzysztof Adamkiewicz, Brian B. Moser, Stanislav Frolov, Tobias Christian Nauen et al.CVPR 2026 · 8 citations
- Vision Transformer Neural Architecture Search for Out-of-Distribution Generalization: Benchmark and InsightsSy-Tuyen Ho, Tuan Van Vo, Somayeh Ebrahimkhani, Ngai-Man CheungNeurIPS 2024 · 5 citations
- Automated Detection of Visual Attribute Reliance with a Self-Reflective AgentChristy Li, Josep López Camuñas, Jake Thomas Touchet, Jacob Andreas et al.NeurIPS 2025 · 2 citations
Builds on27
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal et al.NeurIPS 2020 · 5,249 citations
- Big Self-Supervised Models are Strong Semi-Supervised LearnersTing Chen, Simon Kornblith, Kevin Swersky, Mohammad Norouzi et al.NeurIPS 2020 · 2,611 citations
- An Empirical Study of Training Self-Supervised Vision TransformersXinlei Chen, Saining Xie, Kaiming HeICCV 2021 · 2,340 citations
Related papers
- Neural Redshift: Random Networks are not Random FunctionsDamien Teney, Armand Mihai Nicolicioiu, Valentin Hartmann, Ehsan AbbasnejadCVPR 2024 · 7 citations
- Impact of Aliasing on Generalization in Deep Convolutional NetworksCristina Nader Vasconcelos, Hugo Larochelle, Vincent Dumoulin, Rob Romijnders et al.ICCV 2021 · 41 citations
- Knowledge distillation: A good teacher is patient and consistentLucas Beyer, Xiaohua Zhai, Amélie Royer, Larisa Markeeva et al.CVPR 2022 · 215 citations
- Distributional Robustness Loss for Long-tail LearningDvir Samuel, Gal ChechikICCV 2021 · 128 citations
- Counterfactual Generative NetworksAxel Sauer, Andreas GeigerICLR 2021 · 145 citations
