Spatiotemporal Graph Neural Network based Mask Reconstruction for Video Object Segmentation

Liu, Daizong; Xu, Shuangjie; Liu, Xiao-Yang; Xu, Zichuan; Wei, Wei; Zhou, Pan

Computer Science > Computer Vision and Pattern Recognition

arXiv:2012.05499 (cs)

[Submitted on 10 Dec 2020]

Title:Spatiotemporal Graph Neural Network based Mask Reconstruction for Video Object Segmentation

Authors:Daizong Liu, Shuangjie Xu, Xiao-Yang Liu, Zichuan Xu, Wei Wei, Pan Zhou

View PDF

Abstract:This paper addresses the task of segmenting class-agnostic objects in semi-supervised setting. Although previous detection based methods achieve relatively good performance, these approaches extract the best proposal by a greedy strategy, which may lose the local patch details outside the chosen candidate. In this paper, we propose a novel spatiotemporal graph neural network (STG-Net) to reconstruct more accurate masks for video object segmentation, which captures the local contexts by utilizing all proposals. In the spatial graph, we treat object proposals of a frame as nodes and represent their correlations with an edge weight strategy for mask context aggregation. To capture temporal information from previous frames, we use a memory network to refine the mask of current frame by retrieving historic masks in a temporal graph. The joint use of both local patch details and temporal relationships allow us to better address the challenges such as object occlusion and missing. Without online learning and fine-tuning, our STG-Net achieves state-of-the-art performance on four large benchmarks (DAVIS, YouTube-VOS, SegTrack-v2, and YouTube-Objects), demonstrating the effectiveness of the proposed approach.

Comments:	Accepted by AAAI 2021
Subjects:	Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2012.05499 [cs.CV]
	(or arXiv:2012.05499v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2012.05499

Submission history

From: Daizong Liu [view email]
[v1] Thu, 10 Dec 2020 07:57:44 UTC (11,562 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.CV

< prev | next >

new | recent | 2020-12

Change to browse by:

cs
cs.AI

References & Citations

DBLP - CS Bibliography

listing | bibtex

Daizong Liu
Shuangjie Xu
Xiao-Yang Liu
Zichuan Xu
Wei Wei

…

export BibTeX citation

Computer Science > Computer Vision and Pattern Recognition

Title:Spatiotemporal Graph Neural Network based Mask Reconstruction for Video Object Segmentation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Spatiotemporal Graph Neural Network based Mask Reconstruction for Video Object Segmentation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators