Is Joint Training Better for Deep Auto-Encoders?

Open AccessPosted Content

Is Joint Training Better for Deep Auto-Encoders?

- 06 May 2014 -

TLDR

Joint training of deep autoencoders is investigated and it is found that the usage of regularizations in the joint training scheme is crucial in achieving good performance, and in the supervised setting, joint training also shows superior performance when training deeper models.

Abstract:

Traditionally, when generative models of data are developed via deep architectures, greedy layer-wise pre-training is employed. In a well-trained model, the lower layer of the architecture models the data distribution conditional upon the hidden variables, while the higher layers model the hidden distribution prior. But due to the greedy scheme of the layerwise training technique, the parameters of lower layers are fixed when training higher layers. This makes it extremely challenging for the model to learn the hidden distribution prior, which in turn leads to a suboptimal model for the data distribution. We therefore investigate joint training of deep autoencoders, where the architecture is viewed as one stack of two or more single-layer autoencoders. A single global reconstruction objective is jointly optimized, such that the objective for the single autoencoders at each layer acts as a local, layer-level regularizer. We empirically evaluate the performance of this joint training scheme and observe that it not only learns a better data model, but also learns better higher layer representations, which highlights its potential for unsupervised feature learning. In addition, we find that the usage of regularizations in the joint training scheme is crucial in achieving good performance. In the supervised setting, joint training also shows superior performance when training deeper models. The joint training framework can thus provide a platform for investigating more efficient usage of different types of regularizers, especially in light of the growing volumes of available unlabeled data.

Is Joint Training Better for Deep Auto-Encoders?

Citations

Deep learning for visual understanding

Unsupervised Identification of Disease Marker Candidates in Retinal OCT Imaging Data

Review: Deep Learning in Electron Microscopy

Meta-analysis of deep neural networks in remote sensing: A comparative study of mono-temporal classification to support vector machines

Deep learning enabled intelligent fault diagnosis: Overview and applications

References

ImageNet Classification with Deep Convolutional Neural Networks

Learning representations by back-propagating errors

Reducing the Dimensionality of Data with Neural Networks

A fast learning algorithm for deep belief nets

Visualizing and Understanding Convolutional Networks

Related Papers (5)

Reducing the Dimensionality of Data with Neural Networks

Extracting and composing robust features with denoising autoencoders

Stacked Denoising Autoencoders: Learning Useful Representations in a Deep Network with a Local Denoising Criterion

Why Does Unsupervised Pre-training Help Deep Learning?

Contractive Auto-Encoders: Explicit Invariance During Feature Extraction