AI Benchmark: All About Deep Learning on Smartphones in 2019

doi:10.1109/ICCVW.2019.00447

Open AccessProceedings ArticleDOI

AI Benchmark: All About Deep Learning on Smartphones in 2019

Andrey Ignatov, +8 more

- pp 3617-3635

Chats0

TLDR

In this article, the authors evaluate the performance and compare the results of all chipsets from Qualcomm, HiSilicon, Samsung, MediaTek and Unisoc that are providing hardware acceleration for AI inference.

Abstract:

The performance of mobile AI accelerators has been evolving rapidly in the past two years, nearly doubling with each new generation of SoCs. The current 4th generation of mobile NPUs is already approaching the results of CUDA-compatible Nvidia graphics cards presented not long ago, which together with the increased capabilities of mobile deep learning frameworks makes it possible to run complex and deep AI models on mobile devices. In this paper, we evaluate the performance and compare the results of all chipsets from Qualcomm, HiSilicon, Samsung, MediaTek and Unisoc that are providing hardware acceleration for AI inference. We also discuss the recent changes in the Android ML pipeline and provide an overview of the deployment of deep learning models on mobile devices. All numerical results provided in this paper can be found and are regularly updated on the official project website: http://ai-benchmark.com.

Citations

PDF

Open Access

More filters

Proceedings ArticleDOI

MLPerf inference benchmark

Vijay Janapa Reddi, +46 more

TL;DR: This paper presents the benchmarking method for evaluating ML inference systems, MLPerf Inference, and prescribes a set of rules and best practices to ensure comparability across systems with wildly differing architectures.

...read moreread less

Journal ArticleDOI

Pruning and quantization for deep neural network acceleration: A survey

Tailin Liang, +4 more

- 21 Oct 2021 -

Neurocomputing

TL;DR: A survey on two types of network compression: pruning and quantization is provided, which compare current techniques, analyze their strengths and weaknesses, provide guidance for compressing networks, and discuss possible future compression techniques.

...read moreread less

Proceedings ArticleDOI

SPINN: synergistic progressive inference of neural networks over device and cloud

Stefanos Laskaridis, +4 more

TL;DR: SPINN is proposed, a distributed inference system that employs synergistic device-cloud computation together with a progressive inference method to deliver fast and robust CNN inference across diverse settings, and provides robust operation under uncertain connectivity conditions and significant energy savings compared to cloud-centric execution.

...read moreread less

Proceedings ArticleDOI

SPINN: Synergistic Progressive Inference of Neural Networks over Device and Cloud

Stefanos Laskaridis, +4 more

- 14 Aug 2020 -

arXiv: Learning

TL;DR: SPINN as mentioned in this paper proposes a distributed inference system that employs synergistic device-cloud computation together with a progressive inference method to deliver fast and robust CNN inference across diverse settings, and introduces a novel scheduler that co-optimises the early exit policy and the CNN splitting at run time, in order to adapt to dynamic conditions and meet user-defined service-level requirements.

...read moreread less

Collapse

arXiv: Computer Vision and Pattern Recog...

MnasNet: Platform-Aware Neural Architecture Search for Mobile

Mingxing Tan, +6 more

AI Benchmark: All About Deep Learning on Smartphones in 2019

Citations

MLPerf inference benchmark

Pruning and quantization for deep neural network acceleration: A survey

SPINN: synergistic progressive inference of neural networks over device and cloud

Real-Time Quantized Image Super-Resolution on Mobile NPUs, Mobile AI 2021 Challenge: Report

SPINN: Synergistic Progressive Inference of Neural Networks over Device and Cloud

Related Papers (5)

Deep Residual Learning for Image Recognition

MobileNetV2: Inverted Residuals and Linear Bottlenecks

Searching for MobileNetV3

MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications

MnasNet: Platform-Aware Neural Architecture Search for Mobile