Tight Bounds on the Smallest Eigenvalue of the Neural Tangent Kernel for Deep ReLU Networks

Nguyen, Quynh; Mondelli, Marco; Montufar, Guido

Statistics > Machine Learning

arXiv:2012.11654 (stat)

[Submitted on 21 Dec 2020 (v1), last revised 21 Aug 2022 (this version, v5)]

Title:Tight Bounds on the Smallest Eigenvalue of the Neural Tangent Kernel for Deep ReLU Networks

Authors:Quynh Nguyen, Marco Mondelli, Guido Montufar

View PDF

Abstract:A recent line of work has analyzed the theoretical properties of deep neural networks via the Neural Tangent Kernel (NTK). In particular, the smallest eigenvalue of the NTK has been related to the memorization capacity, the global convergence of gradient descent algorithms and the generalization of deep nets. However, existing results either provide bounds in the two-layer setting or assume that the spectrum of the NTK matrices is bounded away from 0 for multi-layer networks. In this paper, we provide tight bounds on the smallest eigenvalue of NTK matrices for deep ReLU nets, both in the limiting case of infinite widths and for finite widths. In the finite-width setting, the network architectures we consider are fairly general: we require the existence of a wide layer with roughly order of $N$ neurons, $N$ being the number of data samples; and the scaling of the remaining layer widths is arbitrary (up to logarithmic factors). To obtain our results, we analyze various quantities of independent interest: we give lower bounds on the smallest singular value of hidden feature matrices, and upper bounds on the Lipschitz constant of input-output feature maps.

Comments:	appeared at ICML 2021, this version corrects a mistake in Lemma 5.4 which also affects Lemma 5.5. These two Lemmas have been edited and the corresponding proofs corrected. All the other results remain untouched
Subjects:	Machine Learning (stat.ML); Machine Learning (cs.LG)
Cite as:	arXiv:2012.11654 [stat.ML]
	(or arXiv:2012.11654v5 [stat.ML] for this version)
	https://doi.org/10.48550/arXiv.2012.11654

Submission history

From: Marco Mondelli [view email]
[v1] Mon, 21 Dec 2020 19:32:17 UTC (44 KB)
[v2] Wed, 23 Dec 2020 20:50:38 UTC (44 KB)
[v3] Tue, 8 Jun 2021 21:27:38 UTC (206 KB)
[v4] Fri, 11 Jun 2021 10:11:24 UTC (332 KB)
[v5] Sun, 21 Aug 2022 13:23:11 UTC (1,381 KB)

Statistics > Machine Learning

Title:Tight Bounds on the Smallest Eigenvalue of the Neural Tangent Kernel for Deep ReLU Networks

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Statistics > Machine Learning

Title:Tight Bounds on the Smallest Eigenvalue of the Neural Tangent Kernel for Deep ReLU Networks

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators