Blazingly Fast Video Object Segmentation with Pixel-Wise Metric Learning

Yuhua Chen, Jordi Pont-Tuset, Alberto Montes, Luc van Gool

2018-04-09CVPR 2018 6Visual Object Tracking Semi-Supervised Video Object Segmentation Metric Learning Segmentation Semantic Segmentation Video Object Segmentation Video Semantic Segmentation Retrieval

Paper PDF

Abstract

This paper tackles the problem of video object segmentation, given some user annotation which indicates the object of interest. The problem is formulated as pixel-wise retrieval in a learned embedding space: we embed pixels of the same object instance into the vicinity of each other, using a fully convolutional network trained by a modified triplet loss as the embedding model. Then the annotated pixels are set as reference and the rest of the pixels are classified using a nearest-neighbor approach. The proposed method supports different kinds of user input such as segmentation mask in the first frame (semi-supervised scenario), or a sparse set of clicked points (interactive scenario). In the semi-supervised scenario, we achieve results competitive with the state of the art but at a fraction of computation cost (275 milliseconds per frame). In the interactive scenario where the user is able to refine their input iteratively, the proposed method provides instant response to each input, and reaches comparable quality to competing methods with much less interaction.

Results

Task	Dataset	Metric	Value	Model
Video	DAVIS 2016	F-measure (Decay)	7.8	PML
Video	DAVIS 2016	F-measure (Mean)	79.3	PML
Video	DAVIS 2016	F-measure (Recall)	93.4	PML
Video	DAVIS 2016	J&F	77.4	PML
Video	DAVIS 2016	Jaccard (Decay)	8.5	PML
Video	DAVIS 2016	Jaccard (Mean)	75.5	PML
Video	DAVIS 2016	Jaccard (Recall)	89.6	PML
Video Object Segmentation	DAVIS 2016	F-measure (Decay)	7.8	PML
Video Object Segmentation	DAVIS 2016	F-measure (Mean)	79.3	PML
Video Object Segmentation	DAVIS 2016	F-measure (Recall)	93.4	PML
Video Object Segmentation	DAVIS 2016	J&F	77.4	PML
Video Object Segmentation	DAVIS 2016	Jaccard (Decay)	8.5	PML
Video Object Segmentation	DAVIS 2016	Jaccard (Mean)	75.5	PML
Video Object Segmentation	DAVIS 2016	Jaccard (Recall)	89.6	PML
Semi-Supervised Video Object Segmentation	DAVIS 2016	F-measure (Decay)	7.8	PML
Semi-Supervised Video Object Segmentation	DAVIS 2016	F-measure (Mean)	79.3	PML
Semi-Supervised Video Object Segmentation	DAVIS 2016	F-measure (Recall)	93.4	PML
Semi-Supervised Video Object Segmentation	DAVIS 2016	J&F	77.4	PML
Semi-Supervised Video Object Segmentation	DAVIS 2016	Jaccard (Decay)	8.5	PML
Semi-Supervised Video Object Segmentation	DAVIS 2016	Jaccard (Mean)	75.5	PML
Semi-Supervised Video Object Segmentation	DAVIS 2016	Jaccard (Recall)	89.6	PML

Blazingly Fast Video Object Segmentation with Pixel-Wise Metric Learning

Abstract

Results

Related Papers

Blazingly Fast Video Object Segmentation with Pixel-Wise Metric Learning

Abstract

Results

Related Papers