Emergence of Shape Bias in Convolutional Neural Networks through Activation Sparsity
Tianqin Li, Ziqi Wen, Yangfan Li, Tai Sing Lee
摘要
Current deep-learning models for object recognition are known to be heavily biased toward texture. In contrast, human visual systems are known to be biased toward shape and structure. What could be the design principles in human visual systems that led to this difference? How could we introduce more shape bias into the deep learning models? In this paper, we report that sparse coding, a ubiquitous principle in the brain, can in itself introduce shape bias into the network. We found that enforcing the sparse coding constraint using a non-differential Top-K operation can lead to the emergence of structural encoding in neurons in convolutional neural networks, resulting in a smooth decomposition of objects into parts and subparts and endowing the networks with shape bias. We demonstrated this emergence of shape bias and its functional benefits for different network structures with various datasets. For object recognition convolutional neural networks, the shape bias leads to greater robustness against style and pattern change distraction. For the image synthesis generative adversary networks, the emerged shape bias leads to more coherent and decomposable structures in the synthesized images. Ablation studies suggest that sparse codes tend to encode structures, whereas the more distributed codes tend to favor texture. Our code is host at the github repository: https://github.com/Crazy-Jack/nips2023_shape_vs_texture
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- MoLe-VLA: Dynamic Layer-skipping Vision Language Action Model via Mixture-of-Layers for Efficient Robot ManipulationRongyu Zhang, Menghang Dong, Yuan Zhang, Liang Heng 等AAAI 2026 · 被引用 56 次
- Visual Anagrams Reveal Hidden Differences in Holistic Shape Processing Across Vision ModelsFenil R. Doshi, Thomas Fel, Talia Konkle, George A. AlvarezNeurIPS 2025 · 被引用 5 次
- Dissecting Generalized Category Discovery: Multiplex Consensus under Self-DeconstructionLuyao Tang, Kunze Huang, Chaoqi Chen, Yuxuan Yuan 等ICCV 2025 · 被引用 3 次
- Feature segregation by signed weights in artificial vision systems and biological modelsGiordano Ramos-Traslosheros, Carlos PonceICLR 2026
- Decomposing the Neurons: Activation Sparsity via Mixture of Experts for Continual Test Time AdaptationRongyu Zhang, Aosong Cheng, Yulin Luo, Gaole Dai 等AAAI 2026
它引用的顶会 Paper7
- Scaling Vision Transformers to 22 Billion ParametersMostafa Dehghani, Josip Djolonga, Basil Mustafa, Piotr Padlewski 等ICML 2023 · 被引用 848 次
- The Origins and Prevalence of Texture Bias in Convolutional Neural NetworksKatherine L. Hermann, Ting Chen, Simon KornblithNeurIPS 2020 · 被引用 369 次
- Learning De-biased Representations with Biased RepresentationsHyojin Bahng, Sanghyuk Chun, Sangdoo Yun, Jaegul Choo 等ICML 2020 · 被引用 332 次
- Towards Faster and Stabilized GAN Training for High-fidelity Few-shot Image SynthesisBingchen Liu, Yizhe Zhu, Kunpeng Song, Ahmed ElgammalICLR 2021 · 被引用 307 次
- Partial success in closing the gap between human and machine visionRobert Geirhos, Kantharaju Narayanappa, Benjamin Mitzkus, Tizian Thieringer 等NeurIPS 2021 · 被引用 304 次
相关 Paper
- Shape or Texture: Understanding Discriminative Features in CNNsMd. Amirul Islam, Matthew Kowal, Patrick Esser, Sen Jia 等ICLR 2021 · 被引用 86 次
- Does enhanced shape bias improve neural network robustness to common corruptions?Chaithanya Kumar Mummadi, Ranjitha Subramaniam, Robin Hutmacher, Julien Vitay 等ICLR 2021 · 被引用 47 次
- Informative Dropout for Robust Representation Learning: A Shape-bias PerspectiveBaifeng Shi, Dinghuai Zhang, Qi Dai, Zhanxing Zhu 等ICML 2020 · 被引用 122 次
- Linear CNNs Discover the Statistical Structure of the Dataset Using Only the Most Dominant FrequenciesHannah Pinson, Joeri Lenaerts, Vincent GinisICML 2023 · 被引用 8 次
- Geometric and Textural Augmentation for Domain Gap ReductionXiao-Chang Liu, Yongliang Yang, Peter HallCVPR 2022 · 被引用 16 次
