On the Learning Dynamics of Two-layer Nonlinear Convolutional Neural Networks

Yu, Bing; Zhang, Junzhao; Zhu, Zhanxing

Computer Science > Machine Learning

arXiv:1905.10157 (cs)

[Submitted on 24 May 2019]

Title:On the Learning Dynamics of Two-layer Nonlinear Convolutional Neural Networks

Authors:Bing Yu, Junzhao Zhang, Zhanxing Zhu

View PDF

Abstract:Convolutional neural networks (CNNs) have achieved remarkable performance in various fields, particularly in the domain of computer vision. However, why this architecture works well remains to be a mystery. In this work we move a small step toward understanding the success of CNNs by investigating the learning dynamics of a two-layer nonlinear convolutional neural network over some specific data distributions. Rather than the typical Gaussian assumption for input data distribution, we consider a more realistic setting that each data point (e.g. image) contains a specific pattern determining its class label. Within this setting, we both theoretically and empirically show that some convolutional filters will learn the key patterns in data and the norm of these filters will dominate during the training process with stochastic gradient descent. And with any high probability, when the number of iterations is sufficiently large, the CNN model could obtain 100% accuracy over the considered data distributions. Our experiments demonstrate that for practical image classification tasks our findings still hold to some extent.

Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:1905.10157 [cs.LG]
	(or arXiv:1905.10157v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1905.10157

Submission history

From: Bing Yu [view email]
[v1] Fri, 24 May 2019 11:33:13 UTC (1,488 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.LG

< prev | next >

new | recent | 2019-05

Change to browse by:

cs
stat
stat.ML

References & Citations

DBLP - CS Bibliography

listing | bibtex

Bing Yu
Junzhao Zhang
Zhanxing Zhu

export BibTeX citation

Computer Science > Machine Learning

Title:On the Learning Dynamics of Two-layer Nonlinear Convolutional Neural Networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:On the Learning Dynamics of Two-layer Nonlinear Convolutional Neural Networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators