Focal Loss for Dense Object Detection

doi:10.1109/TPAMI.2018.2858826

Open AccessJournal ArticleDOI

Focal Loss for Dense Object Detection

Tsung-Yi Lin, +4 more

- 01 Feb 2020 -

IEEE Transactions on Pattern Analysis an...

- Vol. 42, Iss: 2, pp 318-327

Chats0

TLDR

Focal loss as discussed by the authors focuses training on a sparse set of hard examples and prevents the vast number of easy negatives from overwhelming the detector during training, which improves the accuracy of one-stage detectors.

Abstract:

The highest accuracy object detectors to date are based on a two-stage approach popularized by R-CNN, where a classifier is applied to a sparse set of candidate object locations. In contrast, one-stage detectors that are applied over a regular, dense sampling of possible object locations have the potential to be faster and simpler, but have trailed the accuracy of two-stage detectors thus far. In this paper, we investigate why this is the case. We discover that the extreme foreground-background class imbalance encountered during training of dense detectors is the central cause. We propose to address this class imbalance by reshaping the standard cross entropy loss such that it down-weights the loss assigned to well-classified examples. Our novel Focal Loss focuses training on a sparse set of hard examples and prevents the vast number of easy negatives from overwhelming the detector during training. To evaluate the effectiveness of our loss, we design and train a simple dense detector we call RetinaNet. Our results show that when trained with the focal loss, RetinaNet is able to match the speed of previous one-stage detectors while surpassing the accuracy of all existing state-of-the-art two-stage detectors. Code is at: https://github.com/facebookresearch/Detectron .

Citations

PDF

Open Access

More filters

Posted Content

Preparing Lessons: Improve Knowledge Distillation with Better Supervision

Tiancheng Wen, +2 more

- 18 Nov 2019 -

arXiv: Computer Vision and Pattern Recog...

TL;DR: In this paper, the authors introduce two novel approaches, Knowledge Adjustment (KA) and Dynamic Temperature Distillation (DTD), to penalize bad supervision and improve student model, which can get encouraging performance compared with state-of-the-art methods.

...read moreread less

Posted Content

Graph Density-Aware Losses for Novel Compositions in Scene Graph Generation

Boris Knyazev, +5 more

- 17 May 2020 -

arXiv: Computer Vision and Pattern Recog...

TL;DR: A density-normalized edge loss is introduced, which provides more than a two-fold improvement in certain generalization metrics in scene graph generation, and highlights the difficulty of accurately evaluating models using existing metrics, especially on zero/few shots, and introduces a novel weighted metric.

...read moreread less

Book ChapterDOI

Ischemic Stroke Lesion Segmentation in CT Perfusion Scans Using Pyramid Pooling and Focal Loss

S. Mazdak Abulnaga, +1 more

TL;DR: In this article, a fully convolutional neural network was used for segmenting ischemic stroke lesions in CT perfusion images for the ISLES 2018 challenge, which was based on the PSPNet.

...read moreread less

Journal ArticleDOI

Vision-Based Fall Event Detection in Complex Background Using Attention Guided Bi-Directional LSTM

Yong Chen, +4 more

- 04 Sep 2020 -

IEEE Access

TL;DR: Evaluation of the algorithm performances in comparison with other state-of-the-art methods indicates that the proposed design is accurate and robust, which means it is suitable for the task of fall event detection in complex situation.

...read moreread less

Journal ArticleDOI

Using Vehicle Synthesis Generative Adversarial Networks to Improve Vehicle Detection in Remote Sensing Images

Kun Zheng, +4 more

- 04 Sep 2019 -

ISPRS international journal of geo-infor...

TL;DR: A learning method named Vehicle Synthesis Generative Adversarial Networks (VS-GANs) to generate annotated vehicles from remote sensing images that significantly improves the generalization capability of vehicle detectors.

...read moreread less

Collapse

References

PDF

Open Access

More filters

Proceedings ArticleDOI

Deep Residual Learning for Image Recognition

Kaiming He, +3 more

TL;DR: In this article, the authors proposed a residual learning framework to ease the training of networks that are substantially deeper than those used previously, which won the 1st place on the ILSVRC 2015 classification task.

...read moreread less

Proceedings Article

ImageNet Classification with Deep Convolutional Neural Networks

Alex Krizhevsky, +2 more

TL;DR: The state-of-the-art performance of CNNs was achieved by Deep Convolutional Neural Networks (DCNNs) as discussed by the authors, which consists of five convolutional layers, some of which are followed by max-pooling layers, and three fully-connected layers with a final 1000-way softmax.

...read moreread less

Proceedings ArticleDOI

Histograms of oriented gradients for human detection

Navneet Dalal, +1 more

TL;DR: It is shown experimentally that grids of histograms of oriented gradient (HOG) descriptors significantly outperform existing feature sets for human detection, and the influence of each stage of the computation on performance is studied.

...read moreread less

Book ChapterDOI

Microsoft COCO: Common Objects in Context

Tsung-Yi Lin, +7 more

TL;DR: A new dataset with the goal of advancing the state-of-the-art in object recognition by placing the question of object recognition in the context of the broader question of scene understanding by gathering images of complex everyday scenes containing common objects in their natural context.

...read moreread less

Proceedings ArticleDOI

Fully convolutional networks for semantic segmentation

Jonathan Long, +2 more

TL;DR: The key insight is to build “fully convolutional” networks that take input of arbitrary size and produce correspondingly-sized output with efficient inference and learning.

...read moreread less

Collapse

Focal Loss for Dense Object Detection

Citations

Preparing Lessons: Improve Knowledge Distillation with Better Supervision

Graph Density-Aware Losses for Novel Compositions in Scene Graph Generation

Ischemic Stroke Lesion Segmentation in CT Perfusion Scans Using Pyramid Pooling and Focal Loss

Vision-Based Fall Event Detection in Complex Background Using Attention Guided Bi-Directional LSTM

Using Vehicle Synthesis Generative Adversarial Networks to Improve Vehicle Detection in Remote Sensing Images

References

Deep Residual Learning for Image Recognition

ImageNet Classification with Deep Convolutional Neural Networks

Histograms of oriented gradients for human detection

Microsoft COCO: Common Objects in Context

Fully convolutional networks for semantic segmentation

Related Papers (5)

Deep Residual Learning for Image Recognition

You Only Look Once: Unified, Real-Time Object Detection

Microsoft COCO: Common Objects in Context

U-Net: Convolutional Networks for Biomedical Image Segmentation

ImageNet Classification with Deep Convolutional Neural Networks