We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

cs.IT

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Computer Science > Machine Learning

Title: Discrete Deep Feature Extraction: A Theory and New Architectures

Abstract: First steps towards a mathematical theory of deep convolutional neural networks for feature extraction were made---for the continuous-time case---in Mallat, 2012, and Wiatowski and B\"olcskei, 2015. This paper considers the discrete case, introduces new convolutional neural network architectures, and proposes a mathematical framework for their analysis. Specifically, we establish deformation and translation sensitivity results of local and global nature, and we investigate how certain structural properties of the input signal are reflected in the corresponding feature vectors. Our theory applies to general filters and general Lipschitz-continuous non-linearities and pooling operators. Experiments on handwritten digit classification and facial landmark detection---including feature importance evaluation---complement the theoretical findings.
Comments: Proc. of International Conference on Machine Learning (ICML), New York, USA, June 2016, to appear
Subjects: Machine Learning (cs.LG); Computer Vision and Pattern Recognition (cs.CV); Information Theory (cs.IT); Neural and Evolutionary Computing (cs.NE); Machine Learning (stat.ML)
Journal reference: Proc. of International Conference on Machine Learning (ICML), New York, USA, pp. 2149-2158, June 2016
Cite as: arXiv:1605.08283 [cs.LG]
  (or arXiv:1605.08283v1 [cs.LG] for this version)

Submission history

From: Thomas Wiatowski [view email]
[v1] Thu, 26 May 2016 13:55:07 GMT (153kb,D)

Link back to: arXiv, form interface, contact.