Learning Position and Target Consistency for Memory-based Video Object Segmentation - Citegraph

Paper Info

Title
Learning Position and Target Consistency for Memory-based Video Object Segmentation

Abstract
This paper studies the problem of semi-supervised video object segmentation(VOS). Multiple works have shown that memory-based approaches can be effective for video object segmentation. They are mostly based on pixel-level matching, both spatially and temporally. The main shortcoming of memory-based approaches is that they do not take into account the sequential order among frames and do not exploit object-level knowledge from the target. To address this limitation, we propose to Learn position and target Consistency framework for Memory-based video object segmentation, termed as LCM. It applies the memory mechanism to retrieve pixels globally, and meanwhile learns position consistency for more reliable segmentation. The learned location response promotes a better discrimination between target and distractors. Besides, LCM introduces an object-level relationship from the target to maintain target consistency, making LCM more robust to error drifting. Experiments show that our LCM achieves state-of-the-art performance on both DAVIS and Youtube-VOS benchmark. And we rank the 1st in the DAVIS 2020 challenge semi-supervised VOS task.

Year	DOI	Venue
2021	10.1109/CVPR46437.2021.00413	2021 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION, CVPR 2021
DocType	ISSN	Citations
Conference	1063-6919	0
PageRank	References	Authors
0.34	0	6

Authors (6 rows)

Cited by (0 rows)

References (0 rows)

Name	Order	Citations	PageRank
L Hu	1	58	11.75
Peng Zhang	2	0	0.68
Bang Zhang	3	0	1.01
Pan Pan	4	3	4.16
Yinghui Xu	5	172	20.23
Rong Jin	6	0	0.34

1