Amulet: Aggregating Multi-level Convolutional Features for Salient Object Detection

doi:10.1109/ICCV.2017.31

Open AccessProceedings ArticleDOI

Amulet: Aggregating Multi-level Convolutional Features for Salient Object Detection

- pp 202-211

TLDR

Amulet is presented, a generic aggregating multi-level convolutional feature framework for salient object detection that provides accurate salient object labeling and performs favorably against state-of-the-art approaches in terms of near all compared evaluation metrics.

Abstract:

Fully convolutional neural networks (FCNs) have shown outstanding performance in many dense labeling problems. One key pillar of these successes is mining relevant information from features in convolutional layers. However, how to better aggregate multi-level convolutional feature maps for salient object detection is underexplored. In this work, we present Amulet, a generic aggregating multi-level convolutional feature framework for salient object detection. Our framework first integrates multi-level feature maps into multiple resolutions, which simultaneously incorporate coarse semantics and fine details. Then it adaptively learns to combine these feature maps at each resolution and predict saliency maps with the combined features. Finally, the predicted results are efficiently fused to generate the final saliency map. In addition, to achieve accurate boundary inference and semantic enhancement, edge-aware feature maps in low-level layers and the predicted results of low resolution features are recursively embedded into the learning framework. By aggregating multi-level convolutional features in this efficient and flexible manner, the proposed saliency model provides accurate salient object labeling. Comprehensive experiments demonstrate that our method performs favorably against state-of-the-art approaches in terms of near all compared evaluation metrics.

Citations

PDF

Open Access

More filters

Book ChapterDOI

Reverse Attention for Salient Object Detection

Shuhan Chen, +3 more

TL;DR: An accurate yet compact deep network for efficient salient object detection that employs residual learning to learn side-output residual features for saliency refinement, which can be achieved with very limited convolutional parameters while keep accuracy.

...read moreread less

Proceedings ArticleDOI

Shifting More Attention to Video Salient Object Detection

Deng-Ping Fan, +3 more

TL;DR: A visual-attention-consistent Densely Annotated VSOD (DAVSOD) dataset, which contains 226 videos with 23,938 frames that cover diverse realistic-scenes, objects, instances and motions, and a baseline model equipped with a saliency shift- aware convLSTM, which can efficiently capture video saliency dynamics through learning human attention-shift behavior is proposed.

...read moreread less

Posted Content

Salient Object Detection in the Deep Learning Era: An In-Depth Survey

Wenguan Wang, +5 more

- 19 Apr 2019 -

arXiv: Computer Vision and Pattern Recog...

TL;DR: This paper reviews deep SOD algorithms from different perspectives, including network architecture, level of supervision, learning paradigm, and object-/instance-level detection, and looks into the generalization and difficulty of existing SOD datasets.

...read moreread less

Book ChapterDOI

Pyramid Dilated Deeper ConvLSTM for Video Salient Object Detection

Hongmei Song, +4 more

TL;DR: This paper proposes a fast video salient object detection model, based on a novel recurrent network architecture, named Pyramid Dilated Bidirectional ConvLSTM (PDB-ConvL STM), which achieves state-of-the-art results on two popular benchmarks, well demonstrating its superior performance and high applicability.

...read moreread less

Proceedings ArticleDOI

A Bi-Directional Message Passing Model for Salient Object Detection

Lu Zhang, +4 more

TL;DR: This paper proposes a novel bi-directional message passing model to integrate multi-level features for salient object detection, and adopts a Multi-scale Context-aware Feature Extraction Module (MCFEM) for multi- level feature maps to capture rich context information.

...read moreread less

Collapse

References

PDF

Open Access

More filters

Proceedings ArticleDOI

Deep Residual Learning for Image Recognition

Kaiming He, +3 more

TL;DR: In this article, the authors proposed a residual learning framework to ease the training of networks that are substantially deeper than those used previously, which won the 1st place on the ILSVRC 2015 classification task.

...read moreread less

Proceedings Article

Very Deep Convolutional Networks for Large-Scale Image Recognition

Karen Simonyan, +1 more

TL;DR: This work investigates the effect of the convolutional network depth on its accuracy in the large-scale image recognition setting using an architecture with very small convolution filters, which shows that a significant improvement on the prior-art configurations can be achieved by pushing the depth to 16-19 weight layers.

...read moreread less

Proceedings Article

Very Deep Convolutional Networks for Large-Scale Image Recognition

Karen Simonyan, +1 more

TL;DR: In this paper, the authors investigated the effect of the convolutional network depth on its accuracy in the large-scale image recognition setting and showed that a significant improvement on the prior-art configurations can be achieved by pushing the depth to 16-19 layers.

...read moreread less

Book ChapterDOI

U-Net: Convolutional Networks for Biomedical Image Segmentation

Olaf Ronneberger, +2 more

TL;DR: Neber et al. as discussed by the authors proposed a network and training strategy that relies on the strong use of data augmentation to use the available annotated samples more efficiently, which can be trained end-to-end from very few images and outperforms the prior best method (a sliding-window convolutional network) on the ISBI challenge for segmentation of neuronal structures in electron microscopic stacks.

...read moreread less

Proceedings ArticleDOI

Fully convolutional networks for semantic segmentation

Jonathan Long, +2 more

TL;DR: The key insight is to build “fully convolutional” networks that take input of arbitrary size and produce correspondingly-sized output with efficient inference and learning.

...read moreread less

Collapse

Amulet: Aggregating Multi-level Convolutional Features for Salient Object Detection

Citations

Reverse Attention for Salient Object Detection

Shifting More Attention to Video Salient Object Detection

Salient Object Detection in the Deep Learning Era: An In-Depth Survey

Pyramid Dilated Deeper ConvLSTM for Video Salient Object Detection

A Bi-Directional Message Passing Model for Salient Object Detection

References

Deep Residual Learning for Image Recognition

Very Deep Convolutional Networks for Large-Scale Image Recognition

Very Deep Convolutional Networks for Large-Scale Image Recognition

U-Net: Convolutional Networks for Biomedical Image Segmentation

Fully convolutional networks for semantic segmentation

Related Papers (5)

Saliency Detection via Graph-Based Manifold Ranking

Global contrast based salient region detection

Hierarchical Saliency Detection

Deep Residual Learning for Image Recognition

The Secrets of Salient Object Segmentation