Recurrent Dynamic Embedding for Video Object Segmentation

Mingxing Li, Li Hu, Zhiwei Xiong, Bang Zhang, Pan Pan, Dong Liu

2022-05-08CVPR 2022 1Semi-Supervised Video Object Segmentation Semantic Segmentation Video Object Segmentation Video Semantic Segmentation

Paper PDF Code(official)

Abstract

Space-time memory (STM) based video object segmentation (VOS) networks usually keep increasing memory bank every several frames, which shows excellent performance. However, 1) the hardware cannot withstand the ever-increasing memory requirements as the video length increases. 2) Storing lots of information inevitably introduces lots of noise, which is not conducive to reading the most important information from the memory bank. In this paper, we propose a Recurrent Dynamic Embedding (RDE) to build a memory bank of constant size. Specifically, we explicitly generate and update RDE by the proposed Spatio-temporal Aggregation Module (SAM), which exploits the cue of historical information. To avoid error accumulation owing to the recurrent usage of SAM, we propose an unbiased guidance loss during the training stage, which makes SAM more robust in long videos. Moreover, the predicted masks in the memory bank are inaccurate due to the inaccurate network inference, which affects the segmentation of the query frame. To address this problem, we design a novel self-correction strategy so that the network can repair the embeddings of masks with different qualities in the memory bank. Extensive experiments show our method achieves the best tradeoff between performance and speed. Code is available at https://github.com/Limingxing00/RDE-VOS-CVPR2022.

Results

Task	Dataset	Metric	Value	Model
Video	MOSE	F	52.9	RDE
Video	MOSE	J	44.6	RDE
Video	MOSE	J&F	48.8	RDE
Video Object Segmentation	MOSE	F	52.9	RDE
Video Object Segmentation	MOSE	J	44.6	RDE
Video Object Segmentation	MOSE	J&F	48.8	RDE
Semi-Supervised Video Object Segmentation	MOSE	F	52.9	RDE
Semi-Supervised Video Object Segmentation	MOSE	J	44.6	RDE
Semi-Supervised Video Object Segmentation	MOSE	J&F	48.8	RDE

Recurrent Dynamic Embedding for Video Object Segmentation

Abstract

Results

Related Papers

Recurrent Dynamic Embedding for Video Object Segmentation

Abstract

Results

Related Papers