Soft-NMS — Improving Object Detection with One Line of Code

doi:10.1109/ICCV.2017.593

Open AccessProceedings ArticleDOI

Soft-NMS — Improving Object Detection with One Line of Code

- pp 5562-5570

TLDR

Soft-NMS as mentioned in this paper decays the detection scores of all other objects as a continuous function of their overlap with M. As per the design of the algorithm, if an object lies within the predefined overlap threshold, it leads to a miss.

Abstract:

Non-maximum suppression is an integral part of the object detection pipeline. First, it sorts all detection boxes on the basis of their scores. The detection box M with the maximum score is selected and all other detection boxes with a significant overlap (using a pre-defined threshold) with M are suppressed. This process is recursively applied on the remaining boxes. As per the design of the algorithm, if an object lies within the predefined overlap threshold, it leads to a miss. To this end, we propose Soft-NMS, an algorithm which decays the detection scores of all other objects as a continuous function of their overlap with M. Hence, no object is eliminated in this process. Soft-NMS obtains consistent improvements for the coco-style mAP metric on standard datasets like PASCAL VOC2007 (1.7% for both R-FCN and Faster-RCNN) and MS-COCO (1.3% for R-FCN and 1.1% for Faster-RCNN) by just changing the NMS algorithm without any additional hyper-parameters. Using Deformable-RFCN, Soft-NMS improves state-of-the-art in object detection from 39.8% to 40.9% with a single model. Further, the computational complexity of Soft-NMS is the same as traditional NMS and hence it can be efficiently implemented. Since Soft-NMS does not require any extra training and is simple to implement, it can be easily integrated into any object detection pipeline. Code for Soft-NMS is publicly available on GitHub http://bit.ly/2nJLNMu.

Citations

PDF

Open Access

More filters

Proceedings ArticleDOI

Predicting Animation Skeletons for 3D Articulated Models via Volumetric Nets

Zhan Xu, +3 more

TL;DR: This work presents a learning method for predicting animation skeletons for input 3D models of articulated characters that predicts animation skeletons that are much more similar to the ones created by humans compared to several alternatives and baselines.

...read moreread less

Journal ArticleDOI

An Attention Mechanism-Improved YOLOv7 Object Detection Algorithm for Hemp Duck Count Estimation

Kailin Jiang, +9 more

- 10 Oct 2022 -

Agriculture

TL;DR: The results of the algorithm performance evaluation show that the intelligent hemp duck counting method proposed in this paper is feasible and can promote the development of smart reliable automated duck counting.

...read moreread less

Journal ArticleDOI

An Improved Light-Weight Traffic Sign Recognition Algorithm Based on YOLOv4-Tiny

Lanmei Wang, +4 more

- 01 Jan 2021 -

IEEE Access

TL;DR: Wang et al. as mentioned in this paper proposed an improved light-weight traffic sign recognition algorithm based on YOLOv4-Tiny, which improves the detection recall rate and target positioning accuracy.

...read moreread less

Proceedings ArticleDOI

Detecting 11K Classes: Large Scale Object Detection Without Fine-Grained Bounding Boxes

Hao Yang, +2 more

TL;DR: This paper proposes a semi-supervised large scale fine-grained detection method, which only needs bounding box annotations of a smaller number of coarse- grained classes and image-level labels of large scalefine-grains classes, and can detect all classes at nearly fully-super supervised accuracy.

...read moreread less

Posted Content

Precise Detection in Densely Packed Scenes

Eran Goldman, +6 more

- 01 Apr 2019 -

arXiv: Computer Vision and Pattern Recog...

TL;DR: In this article, the authors propose a deep learning based method for precise object detection in densely packed scenes, where the Jaccard index is used as a quality score and the EM merging unit is used to resolve detection overlap ambiguities.

...read moreread less

Collapse

References

PDF

Open Access

More filters

Proceedings ArticleDOI

Deep Residual Learning for Image Recognition

Kaiming He, +3 more

TL;DR: In this article, the authors proposed a residual learning framework to ease the training of networks that are substantially deeper than those used previously, which won the 1st place on the ILSVRC 2015 classification task.

...read moreread less

Proceedings ArticleDOI

Histograms of oriented gradients for human detection

Navneet Dalal, +1 more

TL;DR: It is shown experimentally that grids of histograms of oriented gradient (HOG) descriptors significantly outperform existing feature sets for human detection, and the influence of each stage of the computation on performance is studied.

...read moreread less

Journal ArticleDOI

A Computational Approach to Edge Detection

John Canny

- 01 Jun 1986 -

IEEE Transactions on Pattern Analysis an...

TL;DR: There is a natural uncertainty principle between detection and localization performance, which are the two main goals, and with this principle a single operator shape is derived which is optimal at any scale.

...read moreread less

Proceedings ArticleDOI

You Only Look Once: Unified, Real-Time Object Detection

Joseph Redmon, +3 more

TL;DR: Compared to state-of-the-art detection systems, YOLO makes more localization errors but is less likely to predict false positives on background, and outperforms other detection methods, including DPM and R-CNN, when generalizing from natural images to other domains like artwork.

...read moreread less

Proceedings ArticleDOI

Rich Feature Hierarchies for Accurate Object Detection and Semantic Segmentation

Ross Girshick, +3 more

TL;DR: RCNN as discussed by the authors combines CNNs with bottom-up region proposals to localize and segment objects, and when labeled training data is scarce, supervised pre-training for an auxiliary task, followed by domain-specific fine-tuning, yields a significant performance boost.

...read moreread less

Collapse

Soft-NMS — Improving Object Detection with One Line of Code

Citations

Predicting Animation Skeletons for 3D Articulated Models via Volumetric Nets

An Attention Mechanism-Improved YOLOv7 Object Detection Algorithm for Hemp Duck Count Estimation

An Improved Light-Weight Traffic Sign Recognition Algorithm Based on YOLOv4-Tiny

Detecting 11K Classes: Large Scale Object Detection Without Fine-Grained Bounding Boxes

Precise Detection in Densely Packed Scenes

References

Deep Residual Learning for Image Recognition

Histograms of oriented gradients for human detection

A Computational Approach to Edge Detection

You Only Look Once: Unified, Real-Time Object Detection

Rich Feature Hierarchies for Accurate Object Detection and Semantic Segmentation

Related Papers (5)

SSD: Single Shot MultiBox Detector

Deep Residual Learning for Image Recognition

Feature Pyramid Networks for Object Detection

You Only Look Once: Unified, Real-Time Object Detection

Fast R-CNN