A model of double descent for high-dimensional binary linear classification

Zeyu Deng; Abla Kammoun; Christos Thrampoulidis

doi:10.1093/imaiai/iaab002

Back

A model of double descent for high-dimensional binary linear classification

Journal article

Peer reviewed

A model of double descent for high-dimensional binary linear classification

Zeyu Deng, Abla Kammoun and Christos Thrampoulidis

Information and inference, Vol.11(2), pp.435-495

11/06/2022

DOI: https://doi.org/10.1093/imaiai/iaab002

Abstract

Mathematics

Mathematics, Applied

Physical Sciences

Science & Technology

We consider a model for logistic regression where only a subset of features of size p is used for training a linear classifier over n training samples. The classifier is obtained by running gradient descent on logistic loss. For this model, we investigate the dependence of the classification error on the ratio kappa =p/n. First, building on known deterministic results on the implicit bias of gradient descent, we uncover a phase-transition phenomenon for the case of Gaussian features: the classification error of the gradient descent solution is the same as that of the maximum-likelihood solution when kappa < kappa(star), and that of the support vector machine when kappa > kappa(star), where kappa(star) is a phase-transition threshold. Next, using the convex Gaussian min-max theorem, we sharply characterize the performance of both the maximum-likelihood and the support vector machine solutions. Combining these results, we obtain curves that explicitly characterize the classification error for varying values of kappa. The numerical results validate the theoretical predictions and unveil double-descent phenomena that complement similar recent findings in linear regression settings as well as empirical observations in more complex learning scenarios.

Metrics

1 Record Views

Details

Title: A model of double descent for high-dimensional binary linear classification
Creators - without role: Zeyu Deng - University of California, Santa Barbara
Abla Kammoun - King Abdullah University of Science and Technology
Christos Thrampoulidis - University of British Columbia
Publication Details: Information and inference, Vol.11(2), pp.435-495
Publisher: Oxford Univ Press
Number of pages: 61
Grant note: CCF-2009030 / NSF; National Science Foundation (NSF) King Abdullah University of Science and Technology; King Abdullah University of Science & Technology
Identifiers: 9941616908331
Academic Unit: King Abdullah University of Science & Technology
Language: English
Resource Type: Journal article