We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

cond-mat.dis-nn

Change to browse by:

References & Citations

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Condensed Matter > Disordered Systems and Neural Networks

Title: Clustering of solutions in the symmetric binary perceptron

Abstract: The geometrical features of the (non-convex) loss landscape of neural network models are crucial in ensuring successful optimization and, most importantly, the capability to generalize well. While minimizers' flatness consistently correlates with good generalization, there has been little rigorous work in exploring the condition of existence of such minimizers, even in toy models. Here we consider a simple neural network model, the symmetric perceptron, with binary weights. Phrasing the learning problem as a constraint satisfaction problem, the analogous of a flat minimizer becomes a large and dense cluster of solutions, while the narrowest minimizers are isolated solutions. We perform the first steps toward the rigorous proof of the existence of a dense cluster in certain regimes of the parameters, by computing the first and second moment upper bounds for the existence of pairs of arbitrarily close solutions. Moreover, we present a non rigorous derivation of the same bounds for sets of $y$ solutions at fixed pairwise distances.
Subjects: Disordered Systems and Neural Networks (cond-mat.dis-nn); Machine Learning (stat.ML)
Journal reference: J. Stat. Mech. (2020) 073303
DOI: 10.1088/1742-5468/ab99be
Cite as: arXiv:1911.06756 [cond-mat.dis-nn]
  (or arXiv:1911.06756v3 [cond-mat.dis-nn] for this version)

Submission history

From: Carlo Lucibello [view email]
[v1] Fri, 15 Nov 2019 17:14:07 GMT (2497kb,D)
[v2] Mon, 18 Nov 2019 11:24:38 GMT (733kb,D)
[v3] Mon, 11 May 2020 20:29:19 GMT (1087kb,D)

Link back to: arXiv, form interface, contact.