A Survey of Large-Scale Deep Learning Serving System Optimization: Challenges and Opportunities

Yu, Fuxun; Wang, Di; Shangguan, Longfei; Zhang, Minjia; Tang, Xulong; Liu, Chenchen; Chen, Xiang

Computer Science > Machine Learning

arXiv:2111.14247 (cs)

[Submitted on 28 Nov 2021 (v1), last revised 18 Feb 2022 (this version, v2)]

Title:A Survey of Large-Scale Deep Learning Serving System Optimization: Challenges and Opportunities

Authors:Fuxun Yu, Di Wang, Longfei Shangguan, Minjia Zhang, Xulong Tang, Chenchen Liu, Xiang Chen

View PDF

Abstract:Deep Learning (DL) models have achieved superior performance in many application domains, including vision, language, medical, commercial ads, entertainment, etc. With the fast development, both DL applications and the underlying serving hardware have demonstrated strong scaling trends, i.e., Model Scaling and Compute Scaling, for example, the recent pre-trained model with hundreds of billions of parameters with ~TB level memory consumption, as well as the newest GPU accelerators providing hundreds of TFLOPS. With both scaling trends, new problems and challenges emerge in DL inference serving systems, which gradually trends towards Large-scale Deep learning Serving systems (LDS). This survey aims to summarize and categorize the emerging challenges and optimization opportunities for large-scale deep learning serving systems. By providing a novel taxonomy, summarizing the computing paradigms, and elaborating the recent technique advances, we hope that this survey could shed light on new optimization perspectives and motivate novel works in large-scale deep learning system optimization.

Comments:	10 pages, 7 figures
Subjects:	Machine Learning (cs.LG); Distributed, Parallel, and Cluster Computing (cs.DC)
Cite as:	arXiv:2111.14247 [cs.LG]
	(or arXiv:2111.14247v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2111.14247

Submission history

From: Fuxun Yu [view email]
[v1] Sun, 28 Nov 2021 22:14:10 UTC (5,189 KB)
[v2] Fri, 18 Feb 2022 21:07:02 UTC (5,190 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.LG

< prev | next >

new | recent | 2021-11

Change to browse by:

cs
cs.DC

References & Citations

DBLP - CS Bibliography

listing | bibtex

Fuxun Yu
Di Wang
Minjia Zhang
Chenchen Liu
Xiang Chen

export BibTeX citation

Computer Science > Machine Learning

Title:A Survey of Large-Scale Deep Learning Serving System Optimization: Challenges and Opportunities

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:A Survey of Large-Scale Deep Learning Serving System Optimization: Challenges and Opportunities

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators