The Effects of Approximate Multiplication on Convolutional Neural Networks

Kim, Min Soo; Del Barrio, Alberto A.; Kim, HyunJin; Bagherzadeh, Nader

doi:10.1109/TETC.2021.3050989

Computer Science > Machine Learning

arXiv:2007.10500 (cs)

[Submitted on 20 Jul 2020 (v1), last revised 9 Jan 2021 (this version, v2)]

Title:The Effects of Approximate Multiplication on Convolutional Neural Networks

Authors:Min Soo Kim, Alberto A. Del Barrio, HyunJin Kim, Nader Bagherzadeh

View PDF

Abstract:This paper analyzes the effects of approximate multiplication when performing inferences on deep convolutional neural networks (CNNs). The approximate multiplication can reduce the cost of the underlying circuits so that CNN inferences can be performed more efficiently in hardware accelerators. The study identifies the critical factors in the convolution, fully-connected, and batch normalization layers that allow more accurate CNN predictions despite the errors from approximate multiplication. The same factors also provide an arithmetic explanation of why bfloat16 multiplication performs well on CNNs. The experiments are performed with recognized network architectures to show that the approximate multipliers can produce predictions that are nearly as accurate as the FP32 references, without additional training. For example, the ResNet and Inception-v4 models with Mitch-$w$6 multiplication produces Top-5 errors that are within 0.2% compared to the FP32 references. A brief cost comparison of Mitch-$w$6 against bfloat16 is presented, where a MAC operation saves up to 80% of energy compared to the bfloat16 arithmetic. The most far-reaching contribution of this paper is the analytical justification that multiplications can be approximated while additions need to be exact in CNN MAC operations.

Comments:	12 pages, 11 figures, 4 tables, accepted for publication in the IEEE Transactions on Emerging Topics in Computing
Subjects:	Machine Learning (cs.LG); Computer Vision and Pattern Recognition (cs.CV); Signal Processing (eess.SP)
Cite as:	arXiv:2007.10500 [cs.LG]
	(or arXiv:2007.10500v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2007.10500
Related DOI:	https://doi.org/10.1109/TETC.2021.3050989

Submission history

From: Min Soo Kim [view email]
[v1] Mon, 20 Jul 2020 21:52:41 UTC (3,568 KB)
[v2] Sat, 9 Jan 2021 17:06:41 UTC (4,926 KB)

Computer Science > Machine Learning

Title:The Effects of Approximate Multiplication on Convolutional Neural Networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:The Effects of Approximate Multiplication on Convolutional Neural Networks

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators