Abstract:Spatial attention has been widely used to improve the performance of convolutional neural networks. However, it has certain limitations. In this paper, we propose a new perspective on the effectiveness of spatial attention, which is that the spatial attention mechanism essentially solves the problem of convolutional kernel parameter sharing. However, the information contained in the attention map generated by spatial attention is not sufficient for large-size convolutional kernels. Therefore, we propose a novel attention mechanism called Receptive-Field Attention (RFA). Existing spatial attention, such as Convolutional Block Attention Module (CBAM) and Coordinated Attention (CA) focus only on spatial features, which does not fully address the problem of convolutional kernel parameter sharing. In contrast, RFA not only focuses on the receptive-field spatial feature but also provides effective attention weights for large-size convolutional kernels. The Receptive-Field Attention convolutional operation (RFAConv), developed by RFA, represents a new approach to replace the standard convolution operation. It offers nearly negligible increment of computational cost and parameters, while significantly improving network performance. We conducted a series of experiments on ImageNet-1k, COCO, and VOC datasets to demonstrate the superiority of our approach. Of particular importance, we believe that it is time to shift focus from spatial features to receptive-field spatial features for current spatial attention mechanisms. In this way, we can further improve network performance and achieve even better results. The code and pre-trained models for the relevant tasks can be found at <a class="link-external link-https" href="https://github.com/Liuchen1997/RFAConv" rel="external noopener nofollow">this https URL</a>.

Convolution Tells Where to Look

An Attention Module for Convolutional Neural Networks

HAM: Hybrid Attention Module in Deep Convolutional Neural Networks for Image Classification

Attentive Convolution: Equipping CNNs with RNN-style Attention Mechanisms

A feature-wise attention module based on the difference with surrounding features for convolutional neural networks

Coupled Attention Framework of Convolutional Neural Network Based on Computer Intelligence

Pay Attention to Convolution Filters: Towards Fast and Accurate Fine-Grained Transfer Learning

RFAConv: Innovating Spatial Attention and Standard Convolutional Operation

Convolutional Networks with Dense Connectivity

On the Relationship between Self-Attention and Convolutional Layers

Edge Detection via Fusion Difference Convolution

AFINet: Attentive Feature Integration Networks for Image Classification

RefConv: Re-parameterized Refocusing Convolution for Powerful ConvNets

ECA-Net: Efficient Channel Attention for Deep Convolutional Neural Networks

An Image Classification Method Based on Adaptive Attention Mechanism and Feature Extraction Network

TransConvNet: Perform perceptually relevant driver's visual attention predictions

X-volution: On the unification of convolution and self-attention

Conformer: Local Features Coupling Global Representations for Visual Recognition

Convolutional Neural Network optimization via Channel Reassessment Attention module

A Symmetric Efficient Spatial and Channel Attention (ESCA) Module Based on Convolutional Neural Networks