We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

cs.LG

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Computer Science > Machine Learning

Title: Growing an architecture for a neural network

Abstract: We propose a new kind of automatic architecture search algorithm. The algorithm alternates pruning connections and adding neurons, and it is not restricted to layered architectures only. Here architecture is an arbitrary oriented graph with some weights (along with some biases and an activation function), so there may be no layered structure in such a network. The algorithm minimizes the complexity of staying within a given error. We demonstrate our algorithm on the brightness prediction problem of the next point through the previous points on an image. Our second test problem is the approximation of the bivariate function defining the brightness of a black and white image. Our optimized networks significantly outperform the standard solution for neural network architectures in both cases.
Subjects: Machine Learning (cs.LG); Neural and Evolutionary Computing (cs.NE)
MSC classes: 68T07
Cite as: arXiv:2108.02231 [cs.LG]
  (or arXiv:2108.02231v1 [cs.LG] for this version)

Submission history

From: Ekaterina Shemyakova [view email]
[v1] Wed, 4 Aug 2021 18:17:22 GMT (338kb,D)

Link back to: arXiv, form interface, contact.