Distilling Privileged Knowledge for Anomalous Event Detection From Weakly Labeled Videos.

Tianshan Liu,Kin-Man Lam,Jun Kong

doi:10.1109/tnnls.2023.3263966

Abstract

Weakly supervised video anomaly detection (WS-VAD) aims to identify the snippets involving anomalous events in long untrimmed videos, with solely text video-level binary labels. A typical paradigm among the existing text WS-VAD methods is to employ multiple modalities as inputs, e.g., RGB, optical flow, and audio, as they can provide sufficient discriminative clues that are robust to the diverse, complicated real-world scenes. However, such a pipeline has high reliance on the availability of multiple modalities and is computationally expensive and storage demanding in processing long sequences, which limits its use in some applications. To address this dilemma, we propose a privileged knowledge distillation (KD) framework dedicated to the WS-VAD task, which can maintain the benefits of exploiting additional modalities, while avoiding the need for using multimodal data in the inference phase. We argue that the performance of the privileged KD framework mainly depends on two factors: 1) the effectiveness of the multimodal teacher network and 2) the completeness of the useful information transfer. To obtain a reliable teacher network, we propose a text cross-modal interactive learning strategy and an anomaly normal discrimination loss, which target learning task-specific cross-modal features and encourage the separability of anomalous and normal representations, respectively. Furthermore, we design both representation- and text logits-level distillation loss functions, which force the unimodal student network to distill abundant privileged knowledge from the text well-trained multimodal teacher network, in a snippet-to-video fashion. Extensive experimental results on three public benchmarks demonstrate that the proposed privileged KD framework can train a lightweight yet effective detector, for localizing anomaly events under the supervision of video-level annotations.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Distilling Privileged Knowledge for Anomalous Event Detection From Weakly Labeled Videos.

Abstract

Talk to us

Similar Papers

More From: IEEE transactions on neural networks and learning systems

Lead the way for us

Journal: IEEE transactions on neural networks and learning systems	Publication Date: Sep 1, 2024
Citations: 5

Similar Papers

A deep learning knowledge distillation framework using knee MRI and arthroscopy data for meniscus tear detection.
Mengjie Ying ... Xudong Liu
Frontiers in bioengineering and biotechnology | VOL. 11
Mengjie Ying, et. al.Mengjie Ying ... Xudong Liu
15 Jan 2024
Frontiers in bioengineering and biotechnology | VOL. 11

Privileged Knowledge Distillation for SAR Building Extraction
Eungbean Lee ... Somi Jeong
-
Eungbean Lee, et. al.Eungbean Lee ... Somi Jeong
11 Jul 2021
11 Jul 2021

Learning From Music to Visual Storytelling of Shots: A Deep Interactive Learning Mechanism
Jen-Chun Lin ... Wen-Li Wei
-
Jen-Chun Lin, et. al.Jen-Chun Lin ... Wen-Li Wei
12 Oct 2020
12 Oct 2020

Ensemble Learning of Lightweight Deep Learning Models Using Knowledge Distillation for Image Classification
Jaeyong Kang ... Jeonghwan Gwak
Mathematics | VOL. 8
Jaeyong Kang, et. al.Jaeyong Kang ... Jeonghwan Gwak
24 Sep 2020
Mathematics | VOL. 8

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Distilling Privileged Knowledge for Anomalous Event Detection From Weakly Labeled Videos.

Abstract

Talk to us

Similar Papers

More From: IEEE transactions on neural networks and learning systems