Implicit Bias of Large Depth Networks: a Notion of Rank for Nonlinear Functions

Jacot, Arthur

Full-text links:

Download:

Current browse context:

stat.ML

< prev | next >

new | recent | 2209

Statistics > Machine Learning

Title: Implicit Bias of Large Depth Networks: a Notion of Rank for Nonlinear Functions

Authors: Arthur Jacot

(Submitted on 29 Sep 2022 (v1), last revised 23 Mar 2023 (this version, v4))

Abstract: We show that the representation cost of fully connected neural networks with homogeneous nonlinearities - which describes the implicit bias in function space of networks with $L_2$-regularization or with losses such as the cross-entropy - converges as the depth of the network goes to infinity to a notion of rank over nonlinear functions. We then inquire under which conditions the global minima of the loss recover the `true' rank of the data: we show that for too large depths the global minimum will be approximately rank 1 (underestimating the rank); we then argue that there is a range of depths which grows with the number of datapoints where the true rank is recovered. Finally, we discuss the effect of the rank of a classifier on the topology of the resulting class boundaries and show that autoencoders with optimal nonlinear rank are naturally denoising.

Subjects:	Machine Learning (stat.ML); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:2209.15055 [stat.ML]
	(or arXiv:2209.15055v4 [stat.ML] for this version)

Submission history

From: Arthur Jacot [view email]
[v1] Thu, 29 Sep 2022 18:57:51 GMT (217kb,D)
[v2] Tue, 4 Oct 2022 19:05:19 GMT (217kb,D)
[v3] Thu, 13 Oct 2022 22:59:45 GMT (217kb,D)
[v4] Thu, 23 Mar 2023 18:14:14 GMT (133kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> stat > arXiv:2209.15055

Download:

Current browse context:

Change to browse by:

References & Citations

Bookmark

Statistics > Machine Learning

Title: Implicit Bias of Large Depth Networks: a Notion of Rank for Nonlinear Functions

Submission history