Boundary-sensitive Pre-training for Temporal Localization in Videos

Mengmeng Xu,Bernard Ghanem,Xiatian Zhu,Juan-Manuel Perez-Rua,Brais Martinez,Tao Xiang,Victor Escorcia,Li Zhang

doi:10.1109/iccv48922.2021.00713

Abstract

Many video analysis tasks require temporal localization for the detection of content changes. However, most existing models developed for these tasks are pre-trained on general video action classification tasks. This is due to large scale annotation of temporal boundaries in untrimmed videos being expensive. Therefore, no suitable datasets exist that enable pre-training in a manner sensitive to temporal boundaries. In this paper for the first time, we investigate model pre-training for temporal localization by introducing a novel boundary-sensitive pretext (BSP) task. Instead of relying on costly manual annotations of temporal boundaries, we propose to synthesize temporal boundaries in existing video action classification datasets. By defining different ways of synthesizing boundaries, BSP can then be simply conducted in a self-supervised manner via the classification of the boundary types. This enables the learning of video representations that are much more transferable to downstream temporal localization tasks. Extensive experiments show that the proposed BSP is superior and complementary to the existing action classification-based pre-training counterpart, and achieves new state-of-the-art performance on several temporal localization tasks. Please visit our website for more details https://frostinassiky.github.io/bsp.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Boundary-sensitive Pre-training for Temporal Localization in Videos

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Weakly Supervised Temporal Action Localization Using Deep Metric Learning
Ashraful Islam ... Richard J Radke
-
Ashraful Islam, et. al.Ashraful Islam ... Richard J Radke
01 Mar 2020
01 Mar 2020

Weakly Supervised Temporal Action Localization Through Contrast Based Evaluation Networks.
Ziyi Liu ... Le Wang
IEEE transactions on pattern analysis and machine intelligence | VOL. 44
Ziyi Liu, et. al.Ziyi Liu ... Le Wang
01 Jan 2020
IEEE transactions on pattern analysis and machine intelligence | VOL. 44

Temporal Action Localization Based on Temporal Evolution Model and Multiple Instance Learning
Minglei Yang ... Yan Song
-
Minglei Yang, et. al.Minglei Yang ... Yan Song
11 Dec 2018
11 Dec 2018

Enhancing temporal action localization in an end-to-end network through estimation error incorporation
Mozhgan Mokari ... Khosrow Haj Sadeghi
Image and Vision Computing | VOL. 145
Mozhgan Mokari, et. al.Mozhgan Mokari ... Khosrow Haj Sadeghi
27 Mar 2024
Image and Vision Computing | VOL. 145

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Boundary-sensitive Pre-training for Temporal Localization in Videos

Abstract

Talk to us

Similar Papers