FcaNet: Frequency Channel Attention Networks

Open AccessPosted Content

FcaNet: Frequency Channel Attention Networks

- 22 Dec 2020 -

arXiv: Computer Vision and Pattern Recog...

TLDR

Based on the frequency analysis, the authors mathematically proved that the conventional GAP is a special case of the feature decomposition in the frequency domain and proposed FCANet with novel multi-spectral channel attention.

Abstract:

Attention mechanism, especially channel attention, has gained great success in the computer vision field. Many works focus on how to design efficient channel attention mechanisms while ignoring a fundamental problem, i.e., using global average pooling (GAP) as the unquestionable pre-processing method. In this work, we start from a different view and rethink channel attention using frequency analysis. Based on the frequency analysis, we mathematically prove that the conventional GAP is a special case of the feature decomposition in the frequency domain. With the proof, we naturally generalize the pre-processing of channel attention mechanism in the frequency domain and propose FcaNet with novel multi-spectral channel attention. The proposed method is simple but effective. We can change only one line of code in the calculation to implement our method within existing channel attention methods. Moreover, the proposed method achieves state-of-the-art results compared with other channel attention methods on image classification, object detection, and instance segmentation tasks. Our method could improve by 1.8% in terms of Top-1 accuracy on ImageNet compared with the baseline SENet-50, with the same number of parameters and the same computational cost. Our code and models are publicly available at this https URL

Citations

PDF

Open Access

More filters

Posted Content

Attention Mechanisms in Computer Vision: A Survey.

Meng-Hao Guo, +9 more

- 15 Nov 2021 -

arXiv: Computer Vision and Pattern Recog...

TL;DR: A comprehensive review of attention mechanisms in computer vision can be found in this article, which categorizes them according to approach, such as channel attention, spatial attention, temporal attention and branch attention.

...read moreread less

Proceedings ArticleDOI

NTIRE 2021 Challenge on Perceptual Image Quality Assessment

Jinjin Gu, +49 more

TL;DR: The NTIRE 2021 challenge on perceptual image quality assessment (IQA) as discussed by the authors was held in conjunction with the New Trends in Image Restoration and Enhancement workshop (NTIRE) workshop at CVPR 2021.

...read moreread less

Posted Content

NTIRE 2021 Challenge on Perceptual Image Quality Assessment

Jinjin Gu, +49 more

- 07 May 2021 -

arXiv: Image and Video Processing

TL;DR: The NTIRE 2021 challenge on perceptual image quality assessment (IQA) as mentioned in this paper was held in conjunction with the New Trends in Image Restoration and Enhancement workshop (NTIRE) workshop at CVPR 2021.

...read moreread less

Posted Content

Spatial-Angular Attention Network for Light Field Reconstruction

Gaochang Wu, +3 more

- 05 Jul 2020 -

arXiv: Image and Video Processing

TL;DR: A spatial-angular attention network is proposed to perceive non-local correspondences in the light field, and reconstruct high angular resolution light field in an end-to-end manner with superior performance against sparsely-sampled light fields with Non-Lambertian effects.

...read moreread less

Journal ArticleDOI

Spatial-Angular Attention Network for Light Field Reconstruction

Gaochang Wu, +4 more

- 01 Jan 2021 -

IEEE Transactions on Image Processing

TL;DR: Zhang et al. as mentioned in this paper propose a spatial-angular attention network to perceive non-local correspondences in the light field, and reconstruct high angular resolution light field in an end-to-end manner.

...read moreread less

References

PDF

Open Access

More filters

Proceedings ArticleDOI

Deep Residual Learning for Image Recognition

Kaiming He, +3 more

TL;DR: In this article, the authors proposed a residual learning framework to ease the training of networks that are substantially deeper than those used previously, which won the 1st place on the ILSVRC 2015 classification task.

...read moreread less

Proceedings Article

Attention is All you Need

Ashish Vaswani, +7 more

TL;DR: This paper proposed a simple network architecture based solely on an attention mechanism, dispensing with recurrence and convolutions entirely and achieved state-of-the-art performance on English-to-French translation.

...read moreread less

Journal ArticleDOI

ImageNet Large Scale Visual Recognition Challenge

Olga Russakovsky, +11 more

- 01 Dec 2015 -

International Journal of Computer Vision

TL;DR: The ImageNet Large Scale Visual Recognition Challenge (ILSVRC) as mentioned in this paper is a benchmark in object category classification and detection on hundreds of object categories and millions of images, which has been run annually from 2010 to present, attracting participation from more than fifty institutions.

...read moreread less

Book ChapterDOI

Microsoft COCO: Common Objects in Context

Tsung-Yi Lin, +7 more

TL;DR: A new dataset with the goal of advancing the state-of-the-art in object recognition by placing the question of object recognition in the context of the broader question of scene understanding by gathering images of complex everyday scenes containing common objects in their natural context.

...read moreread less

Proceedings ArticleDOI

Feature Pyramid Networks for Object Detection

Tsung-Yi Lin, +5 more

TL;DR: This paper exploits the inherent multi-scale, pyramidal hierarchy of deep convolutional networks to construct feature pyramids with marginal extra cost and achieves state-of-the-art single-model results on the COCO detection benchmark without bells and whistles.

...read moreread less

Collapse

FcaNet: Frequency Channel Attention Networks

Citations

Attention Mechanisms in Computer Vision: A Survey.

NTIRE 2021 Challenge on Perceptual Image Quality Assessment

NTIRE 2021 Challenge on Perceptual Image Quality Assessment

Spatial-Angular Attention Network for Light Field Reconstruction

Spatial-Angular Attention Network for Light Field Reconstruction

References

Deep Residual Learning for Image Recognition

Attention is All you Need

ImageNet Large Scale Visual Recognition Challenge

Microsoft COCO: Common Objects in Context

Feature Pyramid Networks for Object Detection

Related Papers (5)

Deep Residual Learning for Image Recognition

CBAM: Convolutional Block Attention Module

ImageNet: A large-scale hierarchical image database

ECA-Net: Efficient Channel Attention for Deep Convolutional Neural Networks

Very Deep Convolutional Networks for Large-Scale Image Recognition