Weakly Supervised Temporal Action Localization Using Deep Metric Learning

Ashraful Islam, Richard J. Radke

2020-01-21Action Localization Metric Learning Weakly-supervised Temporal Action Localization Temporal Localization Video Understanding Temporal Action Localization

Paper PDF Code(official)

Abstract

Temporal action localization is an important step towards video understanding. Most current action localization methods depend on untrimmed videos with full temporal annotations of action instances. However, it is expensive and time-consuming to annotate both action labels and temporal boundaries of videos. To this end, we propose a weakly supervised temporal action localization method that only requires video-level action instances as supervision during training. We propose a classification module to generate action labels for each segment in the video, and a deep metric learning module to learn the similarity between different action instances. We jointly optimize a balanced binary cross-entropy loss and a metric loss using a standard backpropagation algorithm. Extensive experiments demonstrate the effectiveness of both of these components in temporal localization. We evaluate our algorithm on two challenging untrimmed video datasets: THUMOS14 and ActivityNet1.2. Our approach improves the current state-of-the-art result for THUMOS14 by 6.5% mAP at IoU threshold 0.5, and achieves competitive performance for ActivityNet1.2.

Results

Task	Dataset	Metric	Value	Model
Video	THUMOS’14	mAP IOU@0.1	62.3	DeepMetricLearner
Video	THUMOS’14	mAP IOU@0.3	46.8	DeepMetricLearner
Video	THUMOS’14	mAP IOU@0.5	29.6	DeepMetricLearner
Video	THUMOS’14	mAP IOU@0.7	9.7	DeepMetricLearner
Video	ActivityNet-1.2	mAP IOU@0.1	60.5	DeepMetricLearner
Video	ActivityNet-1.2	mAP IOU@0.3	48.4	DeepMetricLearner
Video	ActivityNet-1.2	mAP IOU@0.5	35.2	DeepMetricLearner
Video	ActivityNet-1.2	mAP IOU@0.7	16.3	DeepMetricLearner
Temporal Action Localization	THUMOS’14	mAP IOU@0.1	62.3	DeepMetricLearner
Temporal Action Localization	THUMOS’14	mAP IOU@0.3	46.8	DeepMetricLearner
Temporal Action Localization	THUMOS’14	mAP IOU@0.5	29.6	DeepMetricLearner
Temporal Action Localization	THUMOS’14	mAP IOU@0.7	9.7	DeepMetricLearner
Temporal Action Localization	ActivityNet-1.2	mAP IOU@0.1	60.5	DeepMetricLearner
Temporal Action Localization	ActivityNet-1.2	mAP IOU@0.3	48.4	DeepMetricLearner
Temporal Action Localization	ActivityNet-1.2	mAP IOU@0.5	35.2	DeepMetricLearner
Temporal Action Localization	ActivityNet-1.2	mAP IOU@0.7	16.3	DeepMetricLearner
Zero-Shot Learning	THUMOS’14	mAP IOU@0.1	62.3	DeepMetricLearner
Zero-Shot Learning	THUMOS’14	mAP IOU@0.3	46.8	DeepMetricLearner
Zero-Shot Learning	THUMOS’14	mAP IOU@0.5	29.6	DeepMetricLearner
Zero-Shot Learning	THUMOS’14	mAP IOU@0.7	9.7	DeepMetricLearner
Zero-Shot Learning	ActivityNet-1.2	mAP IOU@0.1	60.5	DeepMetricLearner
Zero-Shot Learning	ActivityNet-1.2	mAP IOU@0.3	48.4	DeepMetricLearner
Zero-Shot Learning	ActivityNet-1.2	mAP IOU@0.5	35.2	DeepMetricLearner
Zero-Shot Learning	ActivityNet-1.2	mAP IOU@0.7	16.3	DeepMetricLearner
Action Localization	THUMOS’14	mAP IOU@0.1	62.3	DeepMetricLearner
Action Localization	THUMOS’14	mAP IOU@0.3	46.8	DeepMetricLearner
Action Localization	THUMOS’14	mAP IOU@0.5	29.6	DeepMetricLearner
Action Localization	THUMOS’14	mAP IOU@0.7	9.7	DeepMetricLearner
Action Localization	ActivityNet-1.2	mAP IOU@0.1	60.5	DeepMetricLearner
Action Localization	ActivityNet-1.2	mAP IOU@0.3	48.4	DeepMetricLearner
Action Localization	ActivityNet-1.2	mAP IOU@0.5	35.2	DeepMetricLearner
Action Localization	ActivityNet-1.2	mAP IOU@0.7	16.3	DeepMetricLearner

Weakly Supervised Temporal Action Localization Using Deep Metric Learning

Abstract

Results

Related Papers

Weakly Supervised Temporal Action Localization Using Deep Metric Learning

Abstract

Results

Related Papers