We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

math.OC

Change to browse by:

References & Citations

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Mathematics > Optimization and Control

Title: Stochastic Subspace Descent

Abstract: We present two stochastic descent algorithms that apply to unconstrained optimization and are particularly efficient when the objective function is slow to evaluate and gradients are not easily obtained, as in some PDE-constrained optimization and machine learning problems. The basic algorithm projects the gradient onto a random subspace at each iteration, similar to coordinate descent but without restricting directional derivatives to be along the axes. This algorithm is previously known but we provide new analysis. We also extend the popular SVRG method to this framework but without requiring that the objective function be written as a finite sum. We provide proofs of convergence for our methods under various convexity assumptions and show favorable results when compared to gradient descent and BFGS on non-convex problems from the machine learning and shape optimization literature. We also note that our analysis gives a proof that the iterates of SVRG and several other popular first-order stochastic methods, in their original formulation, converge almost surely to the optimum; to our knowledge, prior to this work the iterates of SVRG had only been known to converge in expectation.
Comments: 34 pages, 7 figures, submitted on 4/1/19 Update: Main document: 24 Pages, Supplementary Material 9 pages
Subjects: Optimization and Control (math.OC)
Cite as: arXiv:1904.01145 [math.OC]
  (or arXiv:1904.01145v2 [math.OC] for this version)

Submission history

From: David Kozak [view email]
[v1] Mon, 1 Apr 2019 23:47:28 GMT (692kb,D)
[v2] Mon, 29 Apr 2019 16:02:13 GMT (324kb,D)

Link back to: arXiv, form interface, contact.