Coordinate Attention for Efficient Mobile Network Design
Qibin Hou, Daquan Zhou, Jiashi Feng
摘要
Recent studies on mobile network design have demonstrated the remarkable effectiveness of channel attention (e.g., the Squeeze-and-Excitation attention) for lifting model performance, but they generally neglect the positional information, which is important for generating spatially selective attention maps. In this paper, we propose a novel attention mechanism for mobile networks by embedding positional information into channel attention, which we call "coordinate attention". Unlike channel attention that transforms a feature tensor to a single feature vector via 2D global pooling, the coordinate attention factorizes channel attention into two 1D feature encoding processes that aggregate features along the two spatial directions, respectively. In this way, long-range dependencies can be captured along one spatial direction and meanwhile precise positional information can be preserved along the other spatial direction. The resulting feature maps are then encoded separately into a pair of direction-aware and position-sensitive attention maps that can be complementarily applied to the input feature map to augment the representations of the objects of interest. Our coordinate attention is simple and can be flexibly plugged into classic mobile networks, such as MobileNetV2, MobileNeXt, and EfficientNet with nearly no computational overhead. Extensive experiments demonstrate that our coordinate attention is not only beneficial to ImageNet classification but more interestingly, behaves better in down-stream tasks, such as object detection and semantic segmentation. Code is available at https://github.com/Andrew-Qibin/ CoordAttention.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper30
- FcaNet: Frequency Channel Attention NetworksZequn Qin, Pengyi Zhang, Fei Wu, Xi LiICCV 2021 · 被引用 1,049 次
- GhostNetV2: Enhance Cheap Operation with Long-Range AttentionYehui Tang, Kai Han, Jianyuan Guo, Chang Xu 等NeurIPS 2022 · 被引用 634 次
- Deep Model ReassemblyXingyi Yang, Daquan Zhou, Songhua Liu, Jingwen Ye 等NeurIPS 2022 · 被引用 162 次
- Weakly-Supervised Camouflaged Object Detection with Scribble AnnotationsRuozhen He, Qihua Dong, Jiaying Lin, Rynson W. H. LauAAAI 2023 · 被引用 126 次
- Feature Distillation Interaction Weighting Network for Lightweight Image Super-resolutionGuangwei Gao, Wenjie Li, Juncheng Li, Fei Wu 等AAAI 2022 · 被引用 113 次
它引用的顶会 Paper6
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le 等ICCV 2019 · 被引用 9,163 次
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang 等ICCV 2019 · 被引用 2,972 次
- Attention Augmented Convolutional NetworksIrwan Bello, Barret Zoph, Quoc Le, Ashish Vaswani 等ICCV 2019 · 被引用 1,149 次
- HBONet: Harmonious Bottleneck on Two Orthogonal DimensionsDuo Li, Aojun Zhou, Anbang YaoICCV 2019 · 被引用 42 次
- Strip Pooling: Rethinking Spatial Pooling for Scene ParsingQibin Hou, Li Zhang, Ming-Ming Cheng, Jiashi FengCVPR 2020
相关 Paper
- Squeeze-and-Attention Networks for Semantic SegmentationZilong Zhong, Zhong Qiu Lin, Rene Bidart, Xiaodan Hu 等CVPR 2020
- MCA: Moment Channel Attention NetworksYangbo Jiang, Zhiwei Jiang, Le Han, Zenan Huang 等AAAI 2024 · 被引用 15 次
- ECA-Net: Efficient Channel Attention for Deep Convolutional Neural NetworksQilong Wang, Banggu Wu, Pengfei Zhu, Peihua Li 等CVPR 2020
- Linear Context Transform BlockDongsheng Ruan, Jun Wen, Nenggan Zheng, Min ZhengAAAI 2020 · 被引用 26 次
- Channelized Axial Attention - considering Channel Relation within Spatial Attention for Semantic SegmentationYe Huang, Di Kang, Wenjing Jia, Liu Liu 等AAAI 2022 · 被引用 44 次
