Refining activation downsampling with SoftPool
Alexandros Stergiou, Ronald Poppe, Grigorios Kalliatakis
Abstract
Convolutional Neural Networks (CNNs) use pooling to decrease the size of activation maps. This process is crucial to increase the receptive fields and to reduce computational requirements of subsequent convolutions. An important feature of the pooling operation is the minimization of information loss, with respect to the initial activation maps, without a significant impact on the computation and memory overhead. To meet these requirements, we propose Soft-Pool: a fast and efficient method for exponentially weighted activation downsampling. Through experiments across a range of architectures and pooling methods, we demonstrate that SoftPool can retain more information in the reduced activation maps. This refined downsampling leads to improvements in a CNN's classification accuracy. Experiments with pooling layer substitutions on ImageNet1K show an increase in accuracy over both original architectures and other pooling methods. We also test SoftPool on video datasets for action recognition. Again, through the direct replacement of pooling layers, we observe consistent performance improvements while computational loads and memory requirements remain limited 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 61345f6f-58c4-499c-88c3-b589982cf45fCited by top-tier papers5
- Holomorphic Equilibrium Propagation Computes Exact Gradients Through Finite Size OscillationsAxel Laborieux, Friedemann ZenkeNeurIPS 2022 · 65 citations
- Transformer-based Entity Typing in Knowledge GraphsZhiwei Hu, Víctor Gutiérrez-Basulto, Zhiliang Xiang, Ru Li et al.EMNLP 2022 · 16 citations
- Improving equilibrium propagation without weight symmetry through Jacobian homeostasisAxel Laborieux, Friedemann ZenkeICLR 2024 · 11 citations
- A General Framework for Robust G-Invariance in G-Equivariant NetworksSophia Sanborn, Nina MiolaneNeurIPS 2023 · 8 citations
- DA-Font: Few-Shot Font Generation via Dual-Attention Hybrid IntegrationWeiran Chen, Guiqian Zhu, Ying Li, Yi Ji et al.ACM MM 2025 · 2 citations
Builds on7
- SlowFast Networks for Video RecognitionChristoph Feichtenhofer, Haoqi Fan, Jitendra Malik, Kaiming HeICCV 2019 · 4,104 citations
- Video Classification With Channel-Separated Convolutional NetworksDu Tran, Heng Wang, Matt Feiszli, Lorenzo TorresaniICCV 2019 · 647 citations
- HACS: Human Action Clips and Segments Dataset for Recognition and Temporal LocalizationHang Zhao, Antonio Torralba, Lorenzo Torresani, Zhicheng YanICCV 2019 · 298 citations
- LIP: Local Importance-Based PoolingZiteng Gao, Limin Wang, Gangshan WuICCV 2019 · 115 citations
- LiftPool: Bidirectional ConvNet PoolingJiaojiao Zhao, Cees G. M. SnoekICLR 2021 · 27 citations
Related papers
- RNNPool: Efficient Non-linear Pooling for RAM Constrained InferenceOindrila Saha, Aditya Kusupati, Harsha Vardhan Simhadri, Manik Varma et al.NeurIPS 2020 · 58 citations
- Global Feature Guided Local PoolingTakumi KobayashiICCV 2019 · 24 citations
- Information Entropy Based Feature Pooling for Convolutional Neural NetworksWeitao Wan, Jiansheng Chen, Tianpeng Li, Yiqing Huang et al.ICCV 2019 · 33 citations
- Mind the Pool: Convolutional Neural Networks Can Overfit Input SizeBilal Alsallakh, David Yan, Narine Kokhlikyan, Vivek Miglani et al.ICLR 2023
- Fast and Accurate Model ScalingPiotr Dollár, Mannat Singh, Ross B. GirshickCVPR 2021
