BQN: Busy-Quiet Net Enabled by Motion Band-Pass Module for Action Recognition.

Guoxi Huang,Adrian G Bors

doi:10.1109/tip.2022.3189810

Abstract

A rich video data representation can be realized by means of spatio-temporal frequency analysis. In this research study we show that a video can be disentangled, following the learning of video characteristics according to their spatio-temporal properties, into two complementary information components, dubbed Busy and Quiet. The Busy information characterizes the boundaries of moving regions, moving objects, or regions of change in movement. Meanwhile, the Quiet information encodes global smooth spatio-temporal structures defined by substantial redundancy. We design a trainable Motion Band-Pass Module (MBPM) for separating Busy and Quiet-defined information, in raw video data. We model a Busy-Quiet Net (BQN) by embedding the MBPM into a two-pathway CNN architecture. The efficiency of BQN is determined by avoiding redundancy in the feature spaces defined by the two pathways. While one pathway processes the Busy features, the other processes Quiet features at lower spatio-temporal resolutions reducing both memory and computational costs. Through experiments we show that the proposed MBPM can be used as a plug-in module in various CNN backbone architectures, significantly boosting their performance. The proposed BQN is shown to outperform many recent video models on Something-Something V1, Kinetics400, UCF101 and HMDB51 datasets.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

BQN: Busy-Quiet Net Enabled by Motion Band-Pass Module for Action Recognition.

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Image Processing

Lead the way for us

Journal: IEEE Transactions on Image Processing	Publication Date: Jan 1, 2022
License type: mit

Similar Papers

Busy-Quiet Video Disentangling for Video Classification
Guoxi Huang ... Adrian G Bors
-
Guoxi Huang, et. al.Guoxi Huang ... Adrian G Bors
01 Jan 2021
01 Jan 2021

EMO-MoviNet: Enhancing Action Recognition in Videos with EvoNorm, Mish Activation, and Optimal Frame Selection for Efficient Mobile Deployment.
Tarique Hussain ... Tanvir Alam
Sensors (Basel, Switzerland) | VOL. 23
Tarique Hussain, et. al.Tarique Hussain ... Tanvir Alam
27 Sep 2023
Sensors (Basel, Switzerland) | VOL. 23

Action Recognition Using Multi-Scale Temporal Shift Module and Temporal Feature Difference Extraction Based on 2D CNN
Kun-Hsuan Wu ... Ching-Te Chiu
Journal of Software Engineering and Applications | VOL. 14
Kun-Hsuan Wu, et. al.Kun-Hsuan Wu ... Ching-Te Chiu
01 Jan 2020
Journal of Software Engineering and Applications | VOL. 14

Two-Level Attention Module Based on Spurious-3D Residual Networks for Human Action Recognition.
Bo Chen ... Fangzhou Meng
Sensors (Basel, Switzerland) | VOL. 23
Bo Chen, et. al.Bo Chen ... Fangzhou Meng
03 Feb 2023
Sensors (Basel, Switzerland) | VOL. 23

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

BQN: Busy-Quiet Net Enabled by Motion Band-Pass Module for Action Recognition.

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Image Processing