PCA-Initialized Deep Neural Networks Applied To Document Image Analysis

Seuret, Mathias; Alberti, Michele; Ingold, Rolf; Liwicki, Marcus

doi:10.1109/ICDAR.2017.148

Full-text links:

Download:

Current browse context:

cs.LG

< prev | next >

new | recent | 1702

Computer Science > Machine Learning

Title: PCA-Initialized Deep Neural Networks Applied To Document Image Analysis

Authors: Mathias Seuret, Michele Alberti, Rolf Ingold, Marcus Liwicki

(Submitted on 1 Feb 2017)

Abstract: In this paper, we present a novel approach for initializing deep neural networks, i.e., by turning PCA into neural layers. Usually, the initialization of the weights of a deep neural network is done in one of the three following ways: 1) with random values, 2) layer-wise, usually as Deep Belief Network or as auto-encoder, and 3) re-use of layers from another network (transfer learning). Therefore, typically, many training epochs are needed before meaningful weights are learned, or a rather similar dataset is required for seeding a fine-tuning of transfer learning. In this paper, we describe how to turn a PCA into an auto-encoder, by generating an encoder layer of the PCA parameters and furthermore adding a decoding layer. We analyze the initialization technique on real documents. First, we show that a PCA-based initialization is quick and leads to a very stable initialization. Furthermore, for the task of layout analysis we investigate the effectiveness of PCA-based initialization and show that it outperforms state-of-the-art random weight initialization methods.

Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Journal reference:	ICDAR 2017
DOI:	10.1109/ICDAR.2017.148
Cite as:	arXiv:1702.00177 [cs.LG]
	(or arXiv:1702.00177v1 [cs.LG] for this version)

Submission history

From: Mathias Seuret [view email]
[v1] Wed, 1 Feb 2017 09:41:52 GMT (8603kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:1702.00177

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Machine Learning

Title: PCA-Initialized Deep Neural Networks Applied To Document Image Analysis

Submission history