Fast Video Salient Object Detection via Spatiotemporal Knowledge Distillation

Tang, Yi; Li, Yuanman; Zou, Wenbin

Computer Science > Computer Vision and Pattern Recognition

arXiv:2010.10027 (cs)

[Submitted on 20 Oct 2020 (v1), last revised 17 Mar 2021 (this version, v2)]

Title:Fast Video Salient Object Detection via Spatiotemporal Knowledge Distillation

Authors:Yi Tang, Yuanman Li, Wenbin Zou

View PDF

Abstract:Since the wide employment of deep learning frameworks in video salient object detection, the accuracy of the recent approaches has made stunning progress. These approaches mainly adopt the sequential modules, based on optical flow or recurrent neural network (RNN), to learn robust spatiotemporal features. These modules are effective but significantly increase the computational burden of the corresponding deep models. In this paper, to simplify the network and maintain the accuracy, we present a lightweight network tailored for video salient object detection through the spatiotemporal knowledge distillation. Specifically, in the spatial aspect, we combine a saliency guidance feature embedding structure and spatial knowledge distillation to refine the spatial features. In the temporal aspect, we propose a temporal knowledge distillation strategy, which allows the network to learn the robust temporal features through the infer-frame feature encoding and distilling information from adjacent frames. The experiments on widely used video datasets (e.g., DAVIS, DAVSOD, SegTrack-V2) prove that our approach achieves competitive performance. Furthermore, without the employment of the complex sequential modules, the proposed network can obtain high efficiency with 0.01s per frame.

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2010.10027 [cs.CV]
	(or arXiv:2010.10027v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2010.10027

Submission history

From: Yi Tang [view email]
[v1] Tue, 20 Oct 2020 04:48:36 UTC (2,691 KB)
[v2] Wed, 17 Mar 2021 09:51:51 UTC (7,726 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Fast Video Salient Object Detection via Spatiotemporal Knowledge Distillation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Fast Video Salient Object Detection via Spatiotemporal Knowledge Distillation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators