Gradient-based learning applied to document recognition

Open Access

Gradient-based learning applied to document recognition

Yann LeCun, +7 more

- pp 306-351

Chats0

TLDR

This paper reviews various methods applied to handwritten character recognition and compares them on a standard handwritten digit recognition task, and Convolutional neural networks are shown to outperform all other techniques.

Abstract:

Multilayer neural networks trained with the back-propagation algorithm constitute the best example of a successful gradient based learning technique. Given an appropriate network architecture, gradient-based learning algorithms can be used to synthesize a complex decision surface that can classify high-dimensional patterns, such as handwritten characters, with minimal preprocessing. This paper reviews various methods applied to handwritten character recognition and compares them on a standard handwritten digit recognition task. Convolutional neural networks, which are specifically designed to deal with the variability of 2D shapes, are shown to outperform all other techniques. Real-life document recognition systems are composed of multiple modules including field extraction, segmentation recognition, and language modeling. A new learning paradigm, called graph transformer networks (GTN), allows such multimodule systems to be trained globally using gradient-based methods so as to minimize an overall performance measure. Two systems for online handwriting recognition are described. Experiments demonstrate the advantage of global training, and the flexibility of graph transformer networks. A graph transformer network for reading a bank cheque is also described. It uses convolutional neural network character recognizers combined with global training techniques to provide record accuracy on business and personal cheques. It is deployed commercially and reads several million cheques per day.

Citations

PDF

Open Access

More filters

Proceedings Article

Conditional Random Fields: Probabilistic Models for Segmenting and Labeling Sequence Data

John Lafferty, +2 more

TL;DR: This work presents iterative parameter estimation algorithms for conditional random fields and compares the performance of the resulting models to HMMs and MEMMs on synthetic and natural-language data.

...read moreread less

Pattern Recognition and Machine Learning

Christopher M. Bishop

TL;DR: Probability distributions of linear models for regression and classification are given in this article, along with a discussion of combining models and combining models in the context of machine learning and classification.

...read moreread less

Proceedings ArticleDOI

Learning a similarity metric discriminatively, with application to face verification

Sumit Chopra, +2 more

TL;DR: The idea is to learn a function that maps input patterns into a target space such that the L/sub 1/ norm in the target space approximates the "semantic" distance in the input space.

...read moreread less

Proceedings Article

Spectral Networks and Locally Connected Networks on Graphs

Joan Bruna, +3 more

TL;DR: This paper considers possible generalizations of CNNs to signals defined on more general domains without the action of a translation group, and proposes two constructions, one based upon a hierarchical clustering of the domain, and another based on the spectrum of the graph Laplacian.

...read moreread less

Proceedings ArticleDOI

Best practices for convolutional neural networks applied to visual document analysis

Patrice Y. Simard, +2 more

TL;DR: A set of concrete bestpractices that document analysis researchers can use to get good results with neural networks, including a simple "do-it-yourself" implementation of convolution with a flexible architecture suitable for many visual document problems.

...read moreread less

Collapse

Neural Computation

ImageNet: A large-scale hierarchical image database

Jia Deng, +5 more

Gradient-based learning applied to document recognition

Citations

Conditional Random Fields: Probabilistic Models for Segmenting and Labeling Sequence Data

Pattern Recognition and Machine Learning

Learning a similarity metric discriminatively, with application to face verification

Spectral Networks and Locally Connected Networks on Graphs

Best practices for convolutional neural networks applied to visual document analysis

Related Papers (5)

ImageNet Classification with Deep Convolutional Neural Networks

Deep Residual Learning for Image Recognition

Going deeper with convolutions

Long short-term memory

ImageNet: A large-scale hierarchical image database