A Model of Double Descent for High-dimensional Binary Linear Classification

Deng, Zeyu; Kammoun, Abla; Thrampoulidis, Christos

Full-text links:

Download:

Current browse context:

stat.ML

< prev | next >

new | recent | 1911

Statistics > Machine Learning

Title: A Model of Double Descent for High-dimensional Binary Linear Classification

Authors: Zeyu Deng, Abla Kammoun, Christos Thrampoulidis

(Submitted on 13 Nov 2019 (this version), latest version 10 May 2020 (v2))

Abstract: We consider a model for logistic regression where only a subset of features of size $p$ is used for training a linear classifier over $n$ training samples. The classifier is obtained by running gradient-descent (GD) on the logistic-loss. For this model, we investigate the dependence of the generalization error on the overparameterization ratio $\kappa=p/n$. First, building on known deterministic results on convergence properties of the GD, we uncover a phase-transition phenomenon for the case of Gaussian regressors: the generalization error of GD is the same as that of the maximum-likelihood (ML) solution when $\kappa<\kappa_\star$, and that of the max-margin (SVM) solution when $\kappa>\kappa_\star$. Next, using the convex Gaussian min-max theorem (CGMT), we sharply characterize the performance of both the ML and SVM solutions. Combining these results, we obtain curves that explicitly characterize the generalization error of GD for varying values of $\kappa$. The numerical results validate the theoretical predictions and unveil double-descent phenomena that complement similar recent observations in linear regression settings.

Comments:	Short version submitted to ICASSP 2020
Subjects:	Machine Learning (stat.ML); Machine Learning (cs.LG); Signal Processing (eess.SP)
Cite as:	arXiv:1911.05822 [stat.ML]
	(or arXiv:1911.05822v1 [stat.ML] for this version)

Submission history

From: Christos Thrampoulidis [view email]
[v1] Wed, 13 Nov 2019 21:41:38 GMT (272kb)
[v2] Sun, 10 May 2020 19:10:56 GMT (372kb)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> stat > arXiv:1911.05822v1

Download:

Current browse context:

Change to browse by:

References & Citations

Bookmark

Statistics > Machine Learning

Title: A Model of Double Descent for High-dimensional Binary Linear Classification

Submission history