Understanding and Training Deep Diagonal Circulant Neural Networks

Araujo, Alexandre; Negrevergne, Benjamin; Chevaleyre, Yann; Atif, Jamal

Computer Science > Machine Learning

arXiv:1901.10255 (cs)

[Submitted on 29 Jan 2019 (v1), last revised 21 Nov 2019 (this version, v3)]

Title:Understanding and Training Deep Diagonal Circulant Neural Networks

Authors:Alexandre Araujo, Benjamin Negrevergne, Yann Chevaleyre, Jamal Atif

View PDF

Abstract:In this paper, we study deep diagonal circulant neural networks, that is deep neural networks in which weight matrices are the product of diagonal and circulant ones. Besides making a theoretical analysis of their expressivity, we introduced principled techniques for training these models: we devise an initialization scheme and proposed a smart use of non-linearity functions in order to train deep diagonal circulant networks. Furthermore, we show that these networks outperform recently introduced deep networks with other types of structured layers. We conduct a thorough experimental study to compare the performance of deep diagonal circulant networks with state of the art models based on structured matrices and with dense models. We show that our models achieve better accuracy than other structured approaches while required 2x fewer weights as the next best approach. Finally we train deep diagonal circulant networks to build a compact and accurate models on a real world video classification dataset with over 3.8 million training examples.

Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:1901.10255 [cs.LG]
	(or arXiv:1901.10255v3 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1901.10255

Submission history

From: Alexandre Araujo [view email]
[v1] Tue, 29 Jan 2019 12:46:35 UTC (55 KB)
[v2] Tue, 11 Jun 2019 16:53:27 UTC (71 KB)
[v3] Thu, 21 Nov 2019 13:52:19 UTC (87 KB)

Computer Science > Machine Learning

Title:Understanding and Training Deep Diagonal Circulant Neural Networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Understanding and Training Deep Diagonal Circulant Neural Networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators