Spatio-Temporal Adaptive Network With Bidirectional Temporal Difference for Action Recognition

Zhilei Li,Yifu Ding,Yuqing Ma,Xianglong Liu,Rui Wang,Zhiping Shi,Jun Li

doi:10.1109/tcsvt.2023.3250646

Abstract

Action Recognition is a fundamental task in computer vision field, with a wide range of applications in autonomous driving, security monitoring, etc. However, previous action recognition approaches usually suffer from the inappropriate spatio-temporal modeling or high computational consumption (e.g., 3D CNN). In this paper, we propose a novel Spatio-Temporal Adaptive Network (STANet) with bidirectional temporal difference, consisting of a Temporal Adaptive module (TA) and a Spatial Adaptive (SA) module, to sufficiently extract the crucial motion information and model the spatial pivotal appearance information from both forward and backward perspectives, respectively. Specifically, the Temporal Adaptive module uses bidirectional temporal differences to learn valuable motion trends and balance the static semantics and dynamic motion for a certain action during information fusion; while the Spatial Adaptive module uses the bidirectional temporal difference to obtain the spatio-channel attention to stress the discriminative position-relevant and semantic-relevant appearance features. Extensive experiments conducted on widely-used action recognition benchmarks UCF-101, HMDB-51, Something-Something V1, and Kinetics-400 prove the effectiveness of the proposed methods compared to other state-of-the-art approaches.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Spatio-Temporal Adaptive Network With Bidirectional Temporal Difference for Action Recognition

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Circuits and Systems for Video Technology

Lead the way for us

Journal: IEEE Transactions on Circuits and Systems for Video Technology	Publication Date: Sep 1, 2023
Citations: 12

Similar Papers

Multipath Attention and Adaptive Gating Network for Video Action Recognition
Haiping Zhang ... Conghao Ma
Neural Processing Letters | VOL. 56
Haiping Zhang, et. al.Haiping Zhang ... Conghao Ma
27 Mar 2024
Neural Processing Letters | VOL. 56

Action Recognition Using Multi-stream 2D CNN with Deep Learning-Based Temporal Modality
Keonwoo Kang ... Sangwoo Park
-
Keonwoo Kang, et. al.Keonwoo Kang ... Sangwoo Park
06 Jan 2023
06 Jan 2023

An Efficient Lightweight Spatio-temporal Attention Module for Action Recognition
Zhonghua Sun ... Kebin Jia
-
Zhonghua Sun, et. al.Zhonghua Sun ... Kebin Jia
17 Nov 2022
17 Nov 2022

Adaptive spatial modulation for spectrally-efficient MIMO spectrum sharing systems
Zied Bouida ... Ali Ghrayeb
-
Zied Bouida, et. al.Zied Bouida ... Ali Ghrayeb
01 Sep 2014
01 Sep 2014

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Spatio-Temporal Adaptive Network With Bidirectional Temporal Difference for Action Recognition

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Circuits and Systems for Video Technology