Normalize Filters! Classical Wisdom for Deep Vision
Gustavo Pérez, Stella X. Yu
摘要
Classical image filters, such as those for averaging or differencing, are carefully normalized to ensure consistency, interpretability, and to avoid artifacts like intensity shifts, halos, or ringing. In contrast, convolutional filters learned end-to-end in deep networks lack such constraints. Although they may resemble wavelets and blob/edge detectors, they are not normalized in the same or any way. Consequently, when images undergo atmospheric transfer, their responses become distorted, leading to incorrect outcomes. We address this limitation by proposing filter normalization, followed by learnable scaling and shifting, akin to batch normalization. This simple yet effective modification ensures that the filters are atmosphere-equivariant, enabling co-domain symmetry. By integrating classical filtering principles into deep learning (applicable to both convolutional neural networks and convolution-dependent vision transformers), our method achieves significant improvements on artificial and natural intensity variation benchmarks. Our ResNet34 could even outperform CLIP by a large margin. Our analysis reveals that unnormalized filters degrade performance, whereas filter normalization regularizes learning, promotes diversity, and improves robustness and generalization.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper7
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 被引用 4,453 次
- CvT: Introducing Convolutions to Vision TransformersHaiping Wu, Bin Xiao, Noel Codella, Mengchen Liu 等ICCV 2021 · 被引用 2,397 次
- Robust And Interpretable Blind Image Denoising Via Bias-Free Convolutional Neural NetworksSreyas Mohan, Zahra Kadkhodaie, Eero P. Simoncelli, Carlos Fernandez-GrandaICLR 2020 · 被引用 154 次
- Tied Block Convolution: Leaner and Better CNNs with Shared Thinner FiltersXudong Wang, Stella X. YuAAAI 2021 · 被引用 50 次
相关 Paper
- Recognizing Instagram Filtered Images with Feature De-StylizationZhe Wu, Zuxuan Wu, Bharat Singh, Larry S. DavisAAAI 2020 · 被引用 20 次
- The Master Key Filters Hypothesis: Deep Filters Are GeneralZahra Babaiee, Peyman M. Kiasari, Daniela Rus, Radu GrosuAAAI 2025 · 被引用 3 次
- Color Equivariant Convolutional NetworksAttila Lengyel, Ombretta Strafforello, Robert-Jan Bruintjes, Alexander Gielisse 等NeurIPS 2023 · 被引用 15 次
- Improving Equivariance in State-of-the-Art Supervised Depth and Normal PredictorsYuanyi Zhong, Anand Bhattad, Yu-Xiong Wang, David A. ForsythICCV 2023 · 被引用 3 次
- Equivariant Adaptation of Large Pretrained ModelsArnab Kumar Mondal, Siba Smarak Panigrahi, Oumar Kaba, Sai Mudumba 等NeurIPS 2023 · 被引用 49 次
