Eigenvalue Decay Implies Polynomial-Time Learnability for Neural Networks

Goel, Surbhi; Klivans, Adam

Full-text links:

Download:

Current browse context:

cs.LG

< prev | next >

new | recent | 1708

Computer Science > Machine Learning

Title: Eigenvalue Decay Implies Polynomial-Time Learnability for Neural Networks

Authors: Surbhi Goel, Adam Klivans

(Submitted on 11 Aug 2017)

Abstract: We consider the problem of learning function classes computed by neural networks with various activations (e.g. ReLU or Sigmoid), a task believed to be computationally intractable in the worst-case. A major open problem is to understand the minimal assumptions under which these classes admit provably efficient algorithms. In this work we show that a natural distributional assumption corresponding to {\em eigenvalue decay} of the Gram matrix yields polynomial-time algorithms in the non-realizable setting for expressive classes of networks (e.g. feed-forward networks of ReLUs). We make no assumptions on the structure of the network or the labels. Given sufficiently-strong polynomial eigenvalue decay, we obtain {\em fully}-polynomial time algorithms in {\em all} the relevant parameters with respect to square-loss. Milder decay assumptions also lead to improved algorithms. This is the first purely distributional assumption that leads to polynomial-time algorithms for networks of ReLUs, even with one hidden layer. Further, unlike prior distributional assumptions (e.g., the marginal distribution is Gaussian), eigenvalue decay has been observed in practice on common data sets.

Subjects:	Machine Learning (cs.LG); Data Structures and Algorithms (cs.DS)
Cite as:	arXiv:1708.03708 [cs.LG]
	(or arXiv:1708.03708v1 [cs.LG] for this version)

Submission history

From: Surbhi Goel [view email]
[v1] Fri, 11 Aug 2017 21:26:05 GMT (25kb)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:1708.03708

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Machine Learning

Title: Eigenvalue Decay Implies Polynomial-Time Learnability for Neural Networks

Submission history