Generalization in Supervised Learning Through Riemannian Contraction

Kozachkov, Leo; Wensing, Patrick M.; Slotine, Jean-Jacques

Full-text links:

Download:

Current browse context:

cs.LG

< prev | next >

new | recent | 2201

Computer Science > Machine Learning

Title: Generalization in Supervised Learning Through Riemannian Contraction

Authors: Leo Kozachkov, Patrick M. Wensing, Jean-Jacques Slotine

(Submitted on 17 Jan 2022 (v1), last revised 26 Jan 2022 (this version, v2))

Abstract: We prove that Riemannian contraction in a supervised learning setting implies generalization. Specifically, we show that if an optimizer is contracting in some Riemannian metric with rate $\lambda > 0$, it is uniformly algorithmically stable with rate $\mathcal{O}(1/\lambda n)$, where $n$ is the number of labelled examples in the training set. The results hold for stochastic and deterministic optimization, in both continuous and discrete-time, for convex and non-convex loss surfaces. The associated generalization bounds reduce to well-known results in the particular case of gradient descent over convex or strongly convex loss surfaces. They can be shown to be optimal in certain linear settings, such as kernel ridge regression under gradient flow.

Comments:	22 pages, 5, figures
Subjects:	Machine Learning (cs.LG); Dynamical Systems (math.DS)
Cite as:	arXiv:2201.06656 [cs.LG]
	(or arXiv:2201.06656v2 [cs.LG] for this version)

Submission history

From: Leo Kozachkov [view email]
[v1] Mon, 17 Jan 2022 23:08:47 GMT (5209kb,D)
[v2] Wed, 26 Jan 2022 18:05:10 GMT (6277kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2201.06656

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Machine Learning

Title: Generalization in Supervised Learning Through Riemannian Contraction

Submission history