Benign overfitting in leaky ReLU networks with moderate input dimension

Karhadkar, Kedar; George, Erin; Murray, Michael; Montúfar, Guido; Needell, Deanna

Computer Science > Machine Learning

arXiv:2403.06903 (cs)

[Submitted on 11 Mar 2024 (v1), last revised 2 Oct 2024 (this version, v3)]

Title:Benign overfitting in leaky ReLU networks with moderate input dimension

Authors:Kedar Karhadkar, Erin George, Michael Murray, Guido Montúfar, Deanna Needell

View PDF HTML (experimental)

Abstract:The problem of benign overfitting asks whether it is possible for a model to perfectly fit noisy training data and still generalize well. We study benign overfitting in two-layer leaky ReLU networks trained with the hinge loss on a binary classification task. We consider input data that can be decomposed into the sum of a common signal and a random noise component, that lie on subspaces orthogonal to one another. We characterize conditions on the signal to noise ratio (SNR) of the model parameters giving rise to benign versus non-benign (or harmful) overfitting: in particular, if the SNR is high then benign overfitting occurs, conversely if the SNR is low then harmful overfitting occurs. We attribute both benign and non-benign overfitting to an approximate margin maximization property and show that leaky ReLU networks trained on hinge loss with gradient descent (GD) satisfy this property. In contrast to prior work we do not require the training data to be nearly orthogonal. Notably, for input dimension $d$ and training sample size $n$, while results in prior work require $d = \Omega(n^2 \log n)$, here we require only $d = \Omega\left(n\right)$.

Comments:	39 pages
Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:2403.06903 [cs.LG]
	(or arXiv:2403.06903v3 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2403.06903

Submission history

From: Kedar Karhadkar [view email]
[v1] Mon, 11 Mar 2024 16:56:01 UTC (62 KB)
[v2] Tue, 9 Jul 2024 23:20:12 UTC (63 KB)
[v3] Wed, 2 Oct 2024 18:52:14 UTC (63 KB)

Computer Science > Machine Learning

Title:Benign overfitting in leaky ReLU networks with moderate input dimension

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Benign overfitting in leaky ReLU networks with moderate input dimension

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators