Neurocube: a programmable digital neuromorphic architecture with high-density 3D memory

doi:10.1145/3007787.3001178

Journal ArticleDOI

Neurocube: a programmable digital neuromorphic architecture with high-density 3D memory

- Vol. 44, Iss: 3, pp 380-392

TLDR

The basic architecture of the Neurocube is presented and an analysis of the logic tier synthesized in 28nm and 15nm process technologies are presented and the performance is evaluated through the mapping of a Convolutional Neural Network and estimating the subsequent power and performance for both training and inference.

Abstract:

This paper presents a programmable and scalable digital neuromorphic architecture based on 3D high-density memory integrated with logic tier for efficient neural computing. The proposed architecture consists of clusters of processing engines, connected by 2D mesh network as a processing tier, which is integrated in 3D with multiple tiers of DRAM. The PE clusters access multiple memory channels (vaults) in parallel. The operating principle, referred to as the memory centric computing, embeds specialized state-machines within the vault controllers of HMC to drive data into the PE clusters. The paper presents the basic architecture of the Neurocube and an analysis of the logic tier synthesized in 28nm and 15nm process technologies. The performance of the Neurocube is evaluated and illustrated through the mapping of a Convolutional Neural Network and estimating the subsequent power and performance for both training and inference.

Citations

PDF

Open Access

More filters

Posted Content

A Modern Primer on Processing in Memory.

Onur Mutlu, +3 more

- 05 Dec 2020 -

arXiv: Hardware Architecture

TL;DR: This chapter discusses recent research that aims to practically enable computation close to data, an approach called processing-in-memory (PIM).

...read moreread less

Posted Content

Neural Cache: Bit-Serial In-Cache Acceleration of Deep Neural Networks

Charles Eckert, +7 more

- 09 May 2018 -

arXiv: Hardware Architecture

TL;DR: This paper presents the Neural Cache architecture, which re-purposes cache structures to transform them into massively parallel compute units capable of running inferences for Deep Neural Networks, and shows that the proposed architecture can improve inference latency and reduce power consumption.

...read moreread less

Journal ArticleDOI

Processing-in-memory: A workload-driven perspective

Saugata Ghose, +5 more

- 08 Aug 2019 -

Ibm Journal of Research and Development

TL;DR: This article describes the work on systematically identifying opportunities for PIM in real applications and quantifies potential gains for popular emerging applications (e.g., machine learning, data analytics, genome analysis) and describes challenges that remain for the widespread adoption of PIM.

...read moreread less

Proceedings ArticleDOI

Laconic deep learning inference acceleration

Sayeh Sharify, +7 more

TL;DR: By decomposing multiplications down to the bit level, the amount of work needed by multiplications during inference can be potentially reduced by at least 40x across a wide selection of neural networks (8b and 16b).

...read moreread less

Proceedings ArticleDOI

Supporting Very Large Models using Automatic Dataflow Graph Partitioning

Minjie Wang, +2 more

- 24 Jul 2018 -

arXiv: Distributed, Parallel, and Cluste...

TL;DR: Tofu as mentioned in this paper partitions a dataflow graph of fine-grained tensor operators in order to work transparently with a general-purpose deep learning platform like MXNet.

...read moreread less

Collapse

References

PDF

Open Access

More filters

Journal ArticleDOI

Gradient-based learning applied to document recognition

Yann LeCun, +6 more

TL;DR: In this article, a graph transformer network (GTN) is proposed for handwritten character recognition, which can be used to synthesize a complex decision surface that can classify high-dimensional patterns, such as handwritten characters.

...read moreread less

Journal ArticleDOI

Deep learning in neural networks

Jürgen Schmidhuber

- 01 Jan 2015 -

Neural Networks

TL;DR: This historical survey compactly summarizes relevant work, much of it from the previous millennium, review deep supervised learning, unsupervised learning, reinforcement learning & evolutionary computation, and indirect search for short programs encoding deep and large networks.

...read moreread less

Book

Neural Networks And Learning Machines

Simon Haykin

TL;DR: Refocused, revised and renamed to reflect the duality of neural networks and learning machines, this edition recognizes that the subject matter is richer when these topics are studied together.

...read moreread less

Journal ArticleDOI

Cellular neural networks: theory

Leon O. Chua, +1 more

- 01 Oct 1988 -

IEEE Transactions on Circuits and System...

TL;DR: In this article, a class of information processing systems called cellular neural networks (CNNs) are proposed, which consist of a massive aggregate of regularly spaced circuit clones, called cells, which communicate with each other directly through their nearest neighbors.

...read moreread less

Book ChapterDOI

GradientBased Learning Applied to Document Recognition

Simon Haykin, +1 more

TL;DR: Various methods applied to handwritten character recognition are reviewed and compared and Convolutional Neural Networks, that are specifically designed to deal with the variability of 2D shapes, are shown to outperform all other techniques.

...read moreread less

Collapse

Neurocube: a programmable digital neuromorphic architecture with high-density 3D memory

Citations

A Modern Primer on Processing in Memory.

Neural Cache: Bit-Serial In-Cache Acceleration of Deep Neural Networks

Processing-in-memory: A workload-driven perspective

Laconic deep learning inference acceleration

Supporting Very Large Models using Automatic Dataflow Graph Partitioning

References

Gradient-based learning applied to document recognition

Deep learning in neural networks

Neural Networks And Learning Machines

Cellular neural networks: theory

GradientBased Learning Applied to Document Recognition

Related Papers (5)

DaDianNao: A Machine-Learning Supercomputer

ISAAC: a convolutional neural network accelerator with in-situ analog arithmetic in crossbars

DianNao: a small-footprint high-throughput accelerator for ubiquitous machine-learning

EIE: efficient inference engine on compressed deep neural network

In-Datacenter Performance Analysis of a Tensor Processing Unit